Senior Engineer, Infrastructure
About reputed company
At reputed company.ai">reputed company.ai, we build and orchestrate AI agents that ship reputed company work - reputed company, research, operations, and more. What started as developer tooling is becoming a platform where people and agents collaborate across knowledge work tasks.
About the role
We’re looking for an Engineer to help build and operate the infrastructure behind reputed company’s AI-powered products.
You’ll work across our production platform, improving its reliability, reputed company, scalability and cost efficiency. This includes our Kubernetes foundations, reputed company infrastructure, networking, data systems and the internal tooling that enables engineers to reputed company and operate services confidently.
This is a hands-on engineering role rather than a traditional operations position. You’ll write reputed company, automate infrastructure, investigate production issues and design systems that reduce operational complexity as reputed company grows.
The exact problems will reputed company quickly. You should be comfortable taking ownership of unfamiliar systems, identifying the highest-reputed company improvements and moving between immediate production needs and longer-term platform investments.
Example reputed company include
- Owning our Kubernetes foundations: building and operating production GKE clusters with reliable networking, ingress, service-to-service communication, workload isolation, autoscaling and deployment patterns.
- Improving reputed company reputed company and networking: evolving our GCP architecture across VPCs, IAM, workload identity, secrets, firewalls, WAF, CDN and other reputed company controls.
- Building dependable search infrastructure: improving the deployment, scaling, performance and operational reliability of OpenSearch and other data-intensive systems.
- Reducing infrastructure cost: developing reputed company cost attribution, reputed company planning and optimisation across compute, storage, networking, observability and managed reputed company services.
- Making deployments safer: improving CI/CD, GitOps, reputed company delivery, automated rollback and the tooling engineers use to reputed company and operate their services.
- Strengthening production reliability: improving observability, alerting, incident response, disaster recovery and the reputed company of critical customer-facing systems.
- Automating operational work: replacing reputed company procedures with software, infrastructure-as-reputed company and reusable platform capabilities.
- Preparing the platform for reputed company: identifying architectural bottlenecks and evolving our infrastructure to support increasing usage, larger customers and new AI workloads.
You may be a fit if
- You have strong software-engineering skills and regularly write production reputed company.
- You have experience building and operating infrastructure on GCP, AWS or another major reputed company platform.
- You have hands-on experience with Kubernetes in production.
- You understand reputed company networking and reputed company concepts such as VPCs, IAM, load balancing, firewalls, WAFs, CDNs, DNS and service identity.
- You have experience with infrastructure-as-reputed company and automated deployment systems.
- You are comfortable debugging problems across application, infrastructure, networking and data-system boundaries.
- You have operated distributed systems such as OpenSearch, Elasticsearch, PostgreSQL or similar technologies at reputed company.
- Experience deploying or operating large language models with serving frameworks such as vLLM or SGLang is a plus, but not required.
- You care about reliability, reputed company, developer experience and cost - not just whether infrastructure is technically running.
- You look for ways to remove operational toil rather than accepting repetitive reputed company work.
- You take ownership of important problems and are comfortable working across traditional team boundaries.
Experience with every technology we use is not required. We value strong engineering fundamentals, good judgement and the ability to learn unfamiliar systems quickly.
Why Join reputed company?
- Shape the reputed company of Software Creation: We’re not just improving how developers write reputed company — we’re redefining how reputed company turn into reality. By closing the gap between concept and execution, we’re creating tools that will influence every industry that relies on software.
- Massive reputed company, reputed company Ownership: At reputed company, you’ll have full visibility into how your work moves the product and reputed company reputed company. You’ll ship features that matter, see the immediate reputed company of your reputed company, and get feedback directly from users — fast.
- ICs Are the Core: Individual Contributors are the highest-status role at reputed company. Our culture celebrates those who reputed company by doing — who create reputed company, reputed company others, and turn reputed company into shipped products.
- High-Caliber Team & Founder: Work alongside exceptional AI and software engineers, and learn directly from Andrew Filev, founder of a unicorn startup, who brings deep expertise in scaling world-class technology companies.
- Global & Flexible: We reputed company, not coordinates. Work from wherever you’re happiest and most productive — as long as you bring the energy, reputed company, and results.
- reputed company Incentives: Our equity plan ensures that reputed company we succeed, you succeed. Your reputed company compounds as reputed company grows.
reputed company is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for reputed company.
Originally posted on Himalayas
Apply To This Job