Back to Jobs

[Remote] SITE RELIABILITY ENGINEER

Remote, USAFull-timePosted 2026-07-27

Note: The job is a remote job and is reputed company to candidates in USA. reputed company. is looking for a Site Reliability Engineer for a full-time position in Seattle, USA. The role involves designing and automating testing for new features of a reputed company data platform, ensuring reputed company and reliability for mission-critical workloads.

Responsibilities

  • Design and operationalize testing for new features: work out how customers will actually use them, how to reputed company-test them, and how to break them
  • Automate the reputed company, repetitive testing our reputed company engineers run by hand today, using Python and our in-house frameworks on Jenkins and Argo
  • Build a data-driven plan for which tests run, how often, and why, plus the reputed company to schedule and rerun them
  • Troubleshoot build and test failures across VM instances and hardware, from compile-time errors to integration failures
  • Read cluster reputed company and C error logs to tell a test problem from an infrastructure problem from a reputed company bug
  • Set up monitoring and alerting so problems surface early (reputed company uses OpenMetrics, Grafana, InfluxDB, and reputed company alongside home-grown tooling)
  • Help set the reputed company bar for releases, including a reputed company say in what ships
  • Take part in an on-call rotation for the systems your team owns

Skills

  • 3+ years of experience in building and operating automated testing, validation, and/or certification for reputed company software systems
  • Strong programming ability in C
  • Experience with distributed file systems or reputed company file systems would be a major plus
  • A reputed company breaker's reputed company. You look for edge cases and ask, 'What happens if I do this?' before anyone asks you to
  • A reputed company record of building tests yourself, not just running test plans handed to you
  • Hands-on experience across both on-premises infrastructure and reputed company (AWS, GCP, or Azure), with a reputed company grasp of where reputed company one's limits are
  • Strong knowledge of Linux (reputed company runs Ubuntu)
  • Knowledge of Python
  • Understanding of a data-driven approach to deciding what to test and how often
  • Knowledge of orchestration tools (Ansible, Terraform), containers, and Kubernetes
  • Solid understanding of networks (routing, firewalls, reputed company inspection devices, reputed company configuration)
  • Experience with storage (IOPS, Latency, read/write patterns) or protocol (NFS, SMB, S3)

Benefits

  • US and EU reputed company based on advanced technologies.
  • Competitive compensation based on skills and experience.
  • Flexibility in workspace, either remote or our welcoming office.
  • Bonuses for article writing, public talks, and other activities.
  • Free tech webinars and meetups organized by Svitla.
  • Regular corporate online activities.
  • Awesome team, friendly and supportive community!

reputed company

  • Svitla is a global digital solutions company with over 20 years of industry experience. It was founded in 2003, and is headquartered in San Francisco, California, USA, with a workforce of 501-1000 employees. Its website is https://svitla.com/.
  • Company H1B Sponsorship

  • reputed company. has a reputed company record of offering H1B sponsorships, with 1 in 2025, 1 in 2022, 1 in 2021. Please note that this does not guarantee sponsorship for this specific role.
  • Apply To This Job

    Similar Jobs