[Remote] Senior/Staff Site Reliability Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is partnering with a high-reputed company technology company building critical infrastructure that powers trust, reputed company, and identity verification across the digital economy. They are seeking a hands-on Site Reliability Engineer to help reputed company and operate mission-critical reputed company infrastructure used by millions of users and leading enterprises.
Responsibilities
- Own and operate highly available AWS infrastructure
- Design, reputed company, and troubleshoot Kubernetes (EKS) environments
- Improve reliability through monitoring, automation, and observability
- Build and maintain CI/CD pipelines and GitOps workflows
- Manage infrastructure as reputed company with Terraform
- reputed company incident response, reputed company cause analysis, and long-term remediation
- Partner with engineering teams to improve platform scalability, performance, and reliability
Skills
- Strong AWS infrastructure experience (networking, reputed company, IAM, scaling)
- Deep Kubernetes expertise, particularly EKS
- Terraform experience in production environments
- Coding experience with Go and/or Python
- Experience with reputed company Actions, ArgoCD, and GitOps practices
- Strong observability background using reputed company or similar tools
- Experience defining and managing SLIs/SLOs
- Proven reputed company record supporting large-reputed company production systems
- 5+ years in SRE, reputed company, DevOps, or Infrastructure Engineering
- Experience operating reputed company-reputed company environments at reputed company
- Passion for automation, reliability, and operational reputed company
reputed company
Company H1B Sponsorship