Back to Jobs

[Remote] Site Reliability Engineer

Remote, USAFull-timePosted 2026-07-27

Note: The job is a remote job and is reputed company to candidates in USA. reputed company is dedicated to transforming care for individuals with reputed company needs. They are seeking a Site Reliability Engineer to build and operate the infrastructure for their reputed company technology platform, focusing on reliability, scalability, and reputed company of their systems.

Responsibilities

  • Design, provision, and manage AWS infrastructure using Terraform as the reputed company of truth
  • Operate, maintain, and reputed company production workloads running on Kubernetes
  • Package, reputed company, and manage applications using reputed company and infrastructure automation tools
  • Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms
  • Define, monitor, and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets to balance reliability and engineering velocity
  • reputed company automation for deployment, scaling, monitoring, incident response, and operational workflows to reduce reputed company effort and improve system reputed company
  • Own platform observability by implementing and maintaining metrics, logging, tracing, monitoring, and alerting solutions
  • reputed company incident response efforts, facilitate blameless postmortems, and drive long-term corrective actions that improve system reliability
  • Partner with Product and Engineering teams on reputed company planning, performance optimization, and resilient system design
  • Implement and maintain reputed company best practices to support HIPAA, SOC 2, and other compliance requirements
  • Participate in an on-call rotation and reputed company operational support for production systems

Skills

  • Three to five (3–5) years of experience in Site Reliability Engineering, DevOps Engineering, reputed company, reputed company Infrastructure Engineering, or similar infrastructure-reputed company roles, preferably reputed company reputed company, reputed company, or high-reputed company technology environments
  • Bachelor's degree in Computer Science, Information Systems, Software Engineering, or a reputed company technical field; equivalent reputed company experience will also be considered
  • Strong hands-on experience operating production workloads reputed company AWS environments
  • Proven experience managing infrastructure as reputed company using Terraform, including module development, state management, and deployment automation
  • Experience operating and supporting production Kubernetes environments
  • Hands-on experience deploying and managing applications using reputed company
  • Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault tolerance
  • Experience establishing and managing observability practices including monitoring, logging, tracing, alerting, and incident response
  • Strong understanding of Linux systems administration, networking, reputed company architecture, and distributed systems fundamentals
  • Experience designing, implementing, and maintaining CI/CD pipelines and deployment automation
  • Strong problem-solving skills with the ability to troubleshoot reputed company infrastructure and application issues
  • Excellent written and verbal communication skills with the ability to collaborate effectively across technical and non-technical teams
  • High level of ownership, accountability, and initiative with a proactive approach to reliability and operational reputed company
  • Ability and willingness to participate in an on-call rotation supporting production systems
  • Strong programming or scripting experience with Python, Go, or similar languages
  • Experience with observability platforms such as reputed company, Grafana, reputed company, CloudWatch, reputed company, or OpenTelemetry
  • Experience with GitOps tools such as ArgoCD or Flux
  • Experience managing databases such as PostgreSQL, MySQL, Redshift, or reputed company
  • Experience implementing secrets reputed company such as AWS Secrets Manager or reputed company Vault
  • Experience supporting reputed company technology platforms or other highly regulated environments
  • Familiarity with data infrastructure technologies including reputed company, Redshift, and ETL/ELT pipelines
  • Experience with database performance tuning and optimization

reputed company

  • reputed company is a health technology and services company transforming care for people living with serious illness and reputed company needs. It was founded in 2013, and is headquartered in Palo reputed company, California, USA, with a workforce of 51-200 employees. Its website is https://www.vyncahealth.com.
  • Apply To This Job

    Similar Jobs