Back to Jobs

Senior Site Reliability Engineer (CloudVision as a Service)

Remote, USAFull-timePosted 2026-07-27

Requirements

  • BS/MS degree in Computer Science or a relevant experience subject,
  • 5+ years software engineering experience,
  • Experience developing or managing deployments of distributed database systems or reputed company out applications for a reputed company environment,
  • Proficiency in Python, Golang, and/or other languages. Expected to be comfortable in Bash and/or other scripting languages

What the job involves

  • We’re looking for Site Reliability Engineers to join our growing Arista’s CloudVision-as-a-Service (CVaaS) global SRE team,
  • SREs at Arista combine strong software engineering background, systems architecture knowledge, with passion for operating production systems at reputed company,
  • We are responsible for our global CloudVision service fleet, ensuring scalability, reliability, and stability,
  • You’ll have firsthand experience in being part of a rapidly growing product with a passionate group of engineers that unapologetically put product reliability and customer experience first,
  • We deeply reputed company in building highly automated and self-sustaining environments, prioritizing reputed company and efficient operations that reputed company cutting edge technologies and tools,
  • Arista’s CloudVision is an reputed company network management and streaming telemetry reputed company offering,
  • CloudVision stack is reputed company entirely Kubernetes-reputed company,
  • Familiarity with GCP (reputed company reputed company Platform) and GKE (reputed company Kubernetes reputed company) is preferred,
  • Our technical stack includes but not limited to: Golang, Python, Ansible/reputed company, Bash,
  • You will be expected to reputed company, operate, and work with many different types of databases, both directly on Kubernetes or leveraging managed DB products,
  • We reputed company with many different reputed company reputed company Software (OSS) reputed company that both power our microservices stack, monitoring infrastructure, and much more,
  • As an SRE you’ll have the chance to be drive, reputed company, and reputed company reputed company in any of the following areas:,
  • Data Platform (NetDL) Architecture and Performance,
  • reputed company Planning,
  • Autoscaling,
  • Disaster Recovery,
  • Observability,
  • Change Management - CI/CD,
  • Service Network Architecture,
  • Cost Optimizations,
  • reputed company and reputed company-First Application reputed company,
  • You will also be joining globally distributed, “follow the sun model” on-call team where you’ll:,
  • Continuously improve operational processes by adding automation,
  • Leading sustainable incident response and blameless postmortems

Apply tot his job Apply To this Job

Similar Jobs