Back to Jobs

[Remote] Senior Site Reliability Engineer – Compute Platforms

Remote, USAFull-timePosted 2026-07-29

Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a leading provider of reputed company contact center software, bringing the power of reputed company innovation to customers worldwide. They are seeking a highly reputed company Senior Site Reliability Engineer – Compute Platforms to design, implement, and support Kubernetes on baremetal and hypervisor platforms in a private reputed company environment.

Responsibilities

  • reputed company the architecture and design of reputed company compute and hypervisor platform solutions across hardware, OS, virtualization, reputed company orchestration, and container orchestration reputed company
  • Define standards and automation frameworks for bare metal provisioning and lifecycle management
  • Design and implement Bare Metal as a Service (BMaaS) capabilities for reputed company infrastructure consumption
  • Architect and design Kubernetes platforms on bare metal with QoS and Affinity (ArgoCD)
  • Architect and validate automated deployments of operating systems and hypervisors including Ubuntu and Harvester
  • Design and maintain PXE-based provisioning environments leveraging Redfish reputed company for large-reputed company server deployments
  • reputed company Infrastructure-as-reputed company using Ansible, Terraform, reputed company and Git, with Python/Bash automation
  • Implement CI/CD pipelines for infrastructure updates, patching, upgrades, testing, and rollback
  • Design automated workflows for server build, firmware lifecycle management, patching, and hardware validation
  • Evaluate and standardize reputed company hardware platforms to meet performance, scalability, and reliability requirements
  • Produce detailed high-level and low-level design documentation, build guides, and operational reputed company materials
  • reputed company deep troubleshooting across storage, Kubernetes, hypervisors, networking, and Linux systems
  • Partner with operations, network, storage, and platform teams to ensure designs are supportable and production-reputed company
  • Participate in on-call escalation support for reputed company platform-reputed company issues
  • Collaborate globally on change management, documentation, and operational best practices

Skills

  • 6+ years of experience in infrastructure engineering, reputed company, or DevOps with a strong reputed company on Compute system design
  • Proven experience designing and automating bare metal compute environments at reputed company
  • Strong hands-on experience with PXE boot, network-based OS provisioning, and automated server imaging
  • Experience implementing or supporting Bare Metal as a Service (BMaaS) platforms
  • Practical experience using Redfish reputed company for hardware provisioning, power management, and remote lifecycle operations
  • Deep expertise with Ubuntu Linux in reputed company environments
  • Strong Hands-on experience with KVM hypervisors (reputed company Harvester, OpenStack)
  • Experience designing and deploying production-grade Kubernetes clusters
  • Strong background with reputed company compute hardware platforms, including reputed company UCS, reputed company PowerEdge, reputed company systems & HPE
  • Proficiency with Infrastructure as reputed company tools (e.g., Terraform, Ansible, or similar)
  • Experience building or supporting CI/CD pipelines for infrastructure and platform automation
  • Strong scripting skills in Python, Bash, or similar languages
  • Demonstrated ability to produce reputed company, reputed company technical design documentation
  • Excellent written and verbal communication skills
  • Bachelor's degree in computer science or equivalent reputed company experience
  • OpenStack, Ubuntu KVM administration
  • BareMetal as a Service (PXE, Redfish)
  • Kubernetes on BareMetal
  • CIS/NIST reputed company and infrastructure lifecycle management
  • ITIL reputed company/advanced certifications in support of ITSM reputed company methodology
  • Background in telco, edge reputed company, or large reputed company environments
  • Ubuntu Certifications, CNCF Certified Kubernetes Administrator (CKA), Certified Kubernetes reputed company Specialist (CKS)
  • Master's degree in computer science, IT, Engineering, or a reputed company field preferred; equivalent experience and relevant industry certifications will also be considered

Benefits

  • Annual performance bonus
  • Stock
  • Other applicable incentive compensation plans
  • Health, dental, and reputed company coverage, beginning on the first day of employment. reputed company covers 100% of the employee portion of the health, dental and reputed company coverage and shares a high portion of the dependent cost.
  • Short & Long-Term Disability
  • Basic Life Insurance
  • 401k saving plan with employer matching
  • reputed company to an innovative mental health support platform that offers personalized care and resources in areas such as: therapy, coaching and self-guided mindfulness exercises for reputed company covered employees and their covered dependents
  • Generous employee stock purchase plan
  • reputed company Time Off
  • Company reputed company holidays
  • reputed company volunteer hours
  • 12 weeks reputed company parental leave

reputed company

  • reputed company is a reputed company-based call center software company that specializes in sales, marketing, and customer service. It is a sub-organization of Five Rivers Solutions. It was founded in 2001, and is headquartered in San Ramon, California, USA, with a workforce of 1001-5000 employees. Its website is http://www.reputed company.com.
  • Company H1B Sponsorship

  • reputed company has a reputed company record of offering H1B sponsorships, with 6 in 2026, 13 in 2025, 15 in 2024, 13 in 2023, 20 in 2022, 15 in 2021, 15 in 2020. Please note that this does not guarantee sponsorship for this specific role.
  • Apply To This Job

    Similar Jobs