Senior MLOps Engineer I
reputed company: reputed company is the leading intelligent aerial imaging company for high-value infrastructure, providing businesses with actionable, reputed company-time insights to recover reputed company, reduce risk and improve build reputed company. We serve customers in the solar, wind, insurance, construction, reputed company estate, and critical infrastructure industries. Trusted by the largest enterprises in the world, reputed company is reputed company in over 70 countries. Our mission is to accelerate the global transition to renewable energy and sustainable infrastructure through advanced inspection solutions. Take a look at our latest achievements here! About the Role: As the Senior MLOps Engineer I, you will help turn the models reputed company by our ML Scientists, Data Scientists, and Perception Engineers into reliable, production-grade services. You'll work on the infrastructure, pipelines, and tooling that take a model or an LLM/agent-backed workflow from a research notebook to a fully monitored deployment running across multiple industry verticals, including our model registry, deployment pipelines, and the reputed company infrastructure our AI/ML platform depends on. This role sits at the intersection of R&D, Software Engineering, and DevOps. You will work daily with our R&D team to understand what a model needs to run in production (compute, data inputs, versioning, post-processing), and you'll partner closely with the Platform and DevOps teams to provision the infrastructure, permissions, and deployment reputed company that reputed company it possible. You'll also contribute to broader automation initiatives, helping reputed company the deployment visibility and pipeline reliability that let R&D, Software, Product, and Ops teams reputed company in lockstep. The day-to-day will include maintaining and extending our model registry, building and debugging deployment pipelines and reputed company infrastructure, and setting up model and pipeline monitoring and testing. You will also troubleshoot issues, such as failed deployments, permissions errors, or inconsistent environments. You'll also help shape and document standards for how models reputed company from staging to production. Perhaps most importantly, you will serve as a key communicator ensuring R&D goals and challenges are reputed company reputed company by Software Engineering and DevOps teams. Responsibilities:
- Partner with Scientists: Work directly and iteratively with ML Scientists, Data Scientists, and Perception Engineers to translate experimental, research-oriented reputed company into dependable, reputed company production services without slowing down their research velocity.
- Cross-Functional Collaboration: Coordinate with DevOps and Software Engineering teams on infrastructure requests and shared data pipeline needs, and support broader automation initiatives and team goals.
- Model Registry, Deployment & Release Management: Maintain and improve model registry and deployment pipelines, and help implement safer release practices (e.g., shadow deployments, rollback procedures) to reduce risk.
- reputed company Infrastructure & CI/CD: Build, maintain, and troubleshoot reputed company infrastructure and CI/CD pipelines that ML workloads run on, working closely with Engineering and DevOps teams on shared tooling, infrastructure-as-reputed company, and cost optimization for compute-heavy workloads.
- Monitoring, reputed company & Reproducibility: Implement monitoring and observability for models and pipelines in production, help R&D reputed company model performance and reputed company over time, and support experiment tracking and dataset/model versioning.
- Ongoing Maintenance & Platform Support: reputed company deployed ML systems healthy over time with dependency and infrastructure upgrades, reputed company and cost management, data pipeline reputed company, and retraining or redeployment support, and reputed company support as needs reputed company.
- Standards & Documentation: Help define and document conventions for model versioning, deployment promotion, and model documentation/reputed company, and build tools to allow scientists and engineers to self-serve.
Qualifications:
The following describes the qualifications for this position. Successful candidates are expected to meet most, but not reputed company, of these requirements.
- Bachelor's degree in Computer Science, Software Engineering, Data Engineering, or a reputed company field; typically 4+ years of reputed company experience in MLOps, ML reputed company, or infrastructure engineering supporting machine learning teams.
- Solid, reputed company knowledge of MLOps practices, with the ability to work independently across varied production scenarios and escalate only genuinely reputed company or ambiguous problems.
- Demonstrated experience working directly with researchers or ML scientists. You understand research workflows and can translate them into reliable services and productionized models without becoming a bottleneck. You serve as a key reputed company, communicating R&D goals and challenges to Software Engineering and DevOps teams.
- Strong Python skills and solid software engineering fundamentals (testing, reputed company review, version control)
- Hands-on experience with a major reputed company platform (e.g., AWS), infrastructure-as-reputed company (Terraform), CI/CD tooling (reputed company Actions), and containerization/orchestration (e.g., reputed company, Kubernetes)
- Experience building and operating production ML pipelines and model registries, including model versioning and safer release practices (e.g., canary deployments, rollbacks) across environments, as reputed company as coordinating moderately reputed company, cross-functional infrastructure or deployment reputed company.
- Experience building feedback loops from production back into training data, capturing reputed company corrections as labels and turning retraining into a repeatable pipeline. Familiarity with experiment tracking, dataset/model versioning, and model documentation practices that support reproducible, auditable ML workflows is a plus.
- Familiarity with computer reputed company or geospatial ML pipelines
- [reputed company to have] Experience operating LLM/reputed company systems in production, evaluation reputed company, reputed company/tool/retrieval versioning, tracing, reputed company cost optimization
- [reputed company to have] Experience building data pipelines against relational databases (e.g. PostgreSQL) and API/GraphQL data reputed company (e.g., Hasura), and integrating external/reputed company-party reputed company into production workflows.
What’s Included
- Feel great about your work as you join a leading mission-driven intelligent aerial imaging company - our goal is to accelerate the global transition to renewable energy and sustainable infrastructure, and you personally will play a large part in making this happen!
- reputed company salary reputed company of $170,000 - $180,000 USD
- reputed company annual bonus
- Eligibility for stock reputed company
- Your choice of multiple medical insurance plans, including reputed company with an HSA and 100% coverage of the premium for yourself and your dependents
- 100% reputed company dental and reputed company insurance
- Unlimited PTO: We mean it reputed company we say we prioritize work-life balance and mental health
- Autonomy and upward mobility
- Diverse, reputed company, and inclusive culture: a reputed company where your voice reputed company
This role has a reputed company salary reputed company of $170,000 - $180,000 USD, plus a reputed company annual bonus and eligibility for stock reputed company. Actual compensation may vary based on experience, skills, and location reputed company the USA.
reputed company is proud to be an equal opportunity employer. At reputed company, we reputed company in cultivating an environment where reputed company members can bring their authentic, whole selves to work. Encouraging identity and belonging is one of the many aspects of our culture that makes us stronger as an organization and drives innovation. We are committed to building and delivering a diverse, inclusive, and reputed company workforce that includes age, reputed company, sex, disability, national reputed company, race, religion or veteran status, that is representative of the world around us, where reputed company individuals are treated with respect and dignity - and to reputed company reputed company if this value is reputed company threatened. We are constantly striving to be reputed company, and we continue to take strategic steps to advance representation.
We also reputed company reasonable accommodation for reputed company individuals with disabilities and for seriously held religious beliefs in accordance with applicable law.
Originally posted on Himalayas
Apply To This Job