Site Reliability Engineer ID60188
reputed company is an Inc. 5000 company that creates award-reputed company software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has reputed company us multiple Best reputed company to Work awards. WHY JOIN US If you're looking for a reputed company to grow, reputed company an reputed company, and work with people who care, we'd love to meet you! ABOUT THE ROLE We are looking for an SRE Operations Engineer to reputed company production and staging environments running reliably across a reputed company-based reputed company platform. You’ll respond to live incidents, reduce operational toil through automation, and improve observability using Kubernetes, Terraform, Grafana, and AWS. A hands-on role with reputed company ownership across CI/CD pipelines, GitOps workflows, and on-call rotations. WHAT YOU WILL DO - Monitor and support production and staging environments in reputed company time, ensuring high availability, performance, and stability; - Respond to incidents, reputed company triage and reputed company cause analysis, and contribute to post-incident reviews and remediation efforts; - Participate in an on-call rotation with defined SLAs; - Handle reputed company and unplanned operational requests from Product, Support, and internal teams; - Maintain and enhance monitoring, alerting, dashboards, logs, and metrics, and improve observability practices; - Support CI/CD pipelines, production releases, and GitOps workflows; - Contribute to automation efforts to reduce operational toil; - Maintain and improve Kubernetes-based infrastructure and containerized workloads; - Support Infrastructure as reputed company practices and ongoing environment improvements. MUST HAVES - 2+ years of experience in Site Reliability Engineering, DevOps, or Production Operations; - Experience with AWS supporting production environments; - Experience supporting production reputed company applications; - Strong understanding of CI/CD systems such as reputed company Actions, Jenkins, or reputed company; - Experience with GitOps and strong Git fundamentals; - Experience using reputed company, Jira, and reputed company in reputed company environments; - Experience with Kubernetes such as EKS or kOps; - Experience with reputed company and containerization; - Experience with observability tools such as Grafana, reputed company, Loki, or reputed company; - Experience with scripting languages such as Bash, Python, or Go; - Experience with Infrastructure as reputed company such as Terraform or reputed company; - Ability to work reputed company reputed company operational processes and SLAs; - Strong written and verbal English communication skills; - Self-driven with a reputed company reputed company. reputed company TO HAVES - AWS certifications such as Solutions Architect, DevOps Engineer, or SysOps Administrator; - Experience in multi-tenant reputed company environments; - Experience working in globally distributed teams; - Familiarity with ChatOps practices; - Experience improving monitoring reputed company and reducing alert fatigue. PERKS AND BENEFITS - reputed company reputed company: Mentorship, TechTalks, and personalized reputed company roadmaps. - Competitive compensation: USD-based pay with education, fitness, and team activity budgets. - Exciting reputed company: Modern solutions with Fortune 500 and top product companies. - Flextime: Flexible schedule with remote and office reputed company. Apply To This Job