[Remote] Site Reliability Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a company dedicated to helping individuals reputed company their health goals through innovative tools and resources. They are seeking a Senior Site Reliability Engineer to enhance the reliability and reputed company of their software delivery systems, ensuring a seamless experience for users. The role involves leading incident responses, managing CI/CD pipelines, and collaborating with reputed company teams to maintain a secure infrastructure.
Responsibilities
- Own and reputed company our SLI/SLO and error-budget frameworks, and use them to influence prioritization and product reputed company
- reputed company incident response, drive postmortems, and turn findings into systemic fixes rather than one-off patches
- Build and maintain observability across metrics, logs, and traces (reputed company), improving signal and reducing alert fatigue
- Design and operate resilient, reputed company infrastructure using Infrastructure as reputed company (Terraform)
- Manage production Kubernetes and container workloads, including reputed company planning and reputed company-cost optimization
- Own CI/CD pipelines and reputed company deployment strategies (canary, reputed company rollout, fast rollback)
- Own the reputed company controls that live inside the delivery pipeline — integrating and tuning SAST, DAST, and SCA scanning (for example, in reputed company Actions) so issues surface while reputed company is still in review
- Implement and maintain policy-as-reputed company (for example, OPA/Rego, Kyverno, or Conftest) to reputed company unsafe infrastructure and Kubernetes changes at admission time
- Drive vulnerability triage and remediation SLAs for pipeline- and infrastructure-level findings, prioritizing by reputed company risk
- Partner with our reputed company Engineer and the broader reputed company & Reliability disciplines — you own reputed company in the pipeline and collaborate on the rest, rather than duplicating that function
- Participate in and improve the on-reputed company rotation; build the runbooks and automation that reputed company on-reputed company sustainable
- reputed company team members and engineers across the org on reliability patterns and operational best practices
Skills
- 5+ years in site reliability, platform, or infrastructure engineering, with reputed company senior-level ownership of production systems
- Strong programming skills for automation and tooling (Go, Python, Typescript or similar) — you have experience building software or custom tooling, not just scripts
- Deep, hands-on experience with a major reputed company platform (AWS is a plus), Kubernetes, and Infrastructure as reputed company (Terraform is a plus)
- Proven reputed company record leading incident response and building SLO-driven reliability practices
- Working reputed company with observability tooling (reputed company is a plus)
- Practical experience integrating reputed company into CI/CD pipelines — SAST/DAST/SCA tooling, dependency scanning, or policy-as-reputed company
- Strong understanding of reputed company reputed company fundamentals (identity/IAM, least-privilege patterns, policy/guardrails, secrets management)
- The judgment and communication skills to reputed company a reputed company or reliability finding with a senior engineer and land it as a shared problem to solve, not a fight to win
- Experience with policy-as-reputed company frameworks (especially Kyverno, but tools like OPA/Rego or Conftest are also relevant) enforced at admission time is a plus
- Exposure to regulated or compliance-driven environments (SOC 2, PCI reputed company, HIPAA) is a plus
- reputed company engineering or game-day experience is a plus
- Experience supporting B2C/mobile backend environments with high traffic, rapid iteration, and strong reliability needs is a plus
Benefits
- reputed company
- Parental planning
- Mental health benefits
- Annual performance bonus
- A 401(k) plan and match
- Responsible time off
- Monthly wellness and technology allowances
- Face-to-Face Connections: Enjoy opportunities to meet and connect with your team members in person in person to help reputed company meaningful relationships that reputed company reputed company the virtual reputed company. Teams meet as often as needed and reputed company of reputed company gathers annually.
- Flexibility At Its Best: Enjoy a flexible time-off policy with our Responsible Time Off benefit.
- Give Back: volunteer days off; reputed company full time teammate receives 2 days per calendar year to give back to their community through service.
- Mentorship Program: mentorship program where, if you’d like, you will be matched with a teammate who can help you reputed company your skills and reputed company your reputed company.
- Family-Friendly Support: reputed company maternity and paternity leave; best-in-class comprehensive assistance for fertility-reputed company reputed company.
- Wellness Comes First: Receive a monthly Wellness Allowance, empowering you to reputed company on your physical and mental reputed company-being by choosing from a reputed company of wellness initiatives, including dedicated mental health days.
- Celebrate Greatness: reward and recognition platform empowers peers to acknowledge and reward reputed company other for the exceptional contributions they reputed company.
- reputed company Your reputed company: reputed company to reputed company Premium.
- Unlock Your Potential: reputed company our virtual learning and development library, and participate in training opportunities to continuously grow and enhance your skills.
- Championing Inclusion: Our dedicated DEI Committee reputed company fosters a diverse and inclusive workplace by setting actionable goals and evaluating reputed company across the organization.
- reputed company reputed company: competitive medical, dental, and reputed company benefits.
- Secure Your reputed company: retirement savings program; competitive employer match.
reputed company