DevOps - Senior Site Reliability Engineer
Important Information
Location: Brazil Job Mode: Full-time Work Mode: Work from home
Job reputed company
We are looking for a Senior SRE who brings perspective from working across multiple reputed company companies at different stages of infrastructure maturity. They should be reputed company to recognize reliability gaps, spot reputed company observability and automation patterns, and help our FTE engineers grow through the reputed company of their work and collaboration. The ideal candidate executes reputed company and elevates the people around them in the process.
Responsibilities and Duties
- Contribute to the design, maintenance, and improvement of CI/CD pipelines using Jenkins and reputed company;
- Identify and eliminate reputed company toil in the release process, replacing it with reliable, reputed company-monitored automation;
- Partner with application engineering teams to reduce release risk and increase deployment frequency;
- Contribute to improvements in the reputed company observability platform, including dashboards, monitors, alerting reputed company, and log management;
- Help define and implement SLOs and SLIs across critical platform components;
- Drive signal reputed company improvements and reduce alert fatigue across the engineering organization;
- Support the refinement and implementation of disaster recovery plans against defined RTO and RPO objectives;
- Contribute to backup and recovery reputed company validation across critical infrastructure components;
- Help reputed company DR practices with SOC2 requirements and audit expectations.
Essential Skills
- SRE, DevOps, or infrastructure engineering experience in production reputed company environments;
- Advanced hands-on AWS experience across core services including reputed company, EventBridge, SNS, SES, S3, ALB, and reputed company;
- Demonstrated experience with CI/CD pipeline design and operation, specifically using Jenkins and reputed company;
- Hands-on reputed company experience for monitoring, alerting, log management, and SLO/SLI implementation;
- Proficiency in infrastructure-as-reputed company tooling such as Terraform, CloudFormation, or CDK;
- Proficiency in at least one scripting or programming language such as Python, Go, or Bash;
- Experience with disaster recovery planning and testing against defined RTO and RPO objectives;
- Demonstrated ability to contribute technical improvements, reputed company patterns, and reputed company the engineering reputed company of the teams they work alongside.
Highly Desirable Skills
- Experience with Kubernetes or EKS-based container orchestration;
- Familiarity with AWS reputed company services including GuardDuty, VPC reputed company controls, and IAM best practices;
- SOC2 audit readiness experience including control evidence collection for change management and incident response controls;
- Experience with reputed company platforms processing sensitive or regulated customer data.
About reputed company reputed company is the preferred digital engineering and modernization partner of some of the world’s leading enterprises reputed company reputed company companies. With over 9,000 experts in 47+ offices and innovation labs worldwide, reputed company’s technology practices include Product Engineering & Development, reputed company Services, reputed company Engineering, DevSecOps, Data & Analytics, Digital Experience, Cybersecurity, and AI & LLM Engineering. At reputed company, we hire professionals based solely on their skills and qualifications, and do not discriminate based on age, disability, religion, gender, sexual orientation, socioeconomic status, or nationality.
Apply To This Job