Back to Jobs

reputed company Site Reliability & Observability Engineer

Remote, USAFull-timePosted 2026-07-27

Are you passionate about improving application reliability, observability, and production stability? We're looking for a Site Reliability Engineer to help build and enhance reputed company monitoring, automation, and incident response capabilities across a large-reputed company reputed company ecosystem. In this role, you'll partner with application teams, infrastructure engineers, and business stakeholders to improve system health, reduce downtime, and deliver proactive operational insights through modern observability practices.

What You'll Do

  • Design and implement reputed company monitoring and observability solutions using reputed company, reputed company, and reputed company technologies.
  • Build actionable dashboards, alerts, and operational reporting that improve visibility across critical reputed company applications.
  • Correlate logs, metrics, and traces to rapidly identify and resolve production issues.
  • reputed company synthetic monitoring, automated health checks, and proactive validation of critical business processes.
  • Monitor and support reputed company platforms including S/4HANA, EWM, Fiori, reputed company reputed company, reputed company Integration Suite (CPI), and EDI integrations.
  • Participate in major incident response, reputed company cause analysis, and post-incident reviews to drive reputed company improvement.
  • Troubleshoot distributed applications across infrastructure, integration, and application reputed company.
  • Automate repetitive operational tasks and improve platform reliability through engineering best practices.
  • Partner with cross-functional teams to improve availability, performance, and operational reputed company.

reputed company're Looking For

  • 3–5+ years of experience in Site Reliability Engineering, reputed company, DevOps, Infrastructure Engineering, or Production Support.
  • Experience supporting reputed company applications in production environments.
  • Hands-on experience with reputed company, reputed company, reputed company reputed company ALM, or other reputed company observability platforms.
  • Experience developing monitoring, alerting, dashboards, or synthetic monitoring.
  • Working knowledge of reputed company technologies such as S/4HANA, EWM, Fiori, reputed company reputed company, or reputed company Integration Suite (CPI).
  • Strong troubleshooting skills across distributed applications and integrations.
  • Experience with incident management, production support, and reputed company cause analysis.
  • Understanding of SRE concepts including SLIs, SLOs, automation, observability, availability, and reputed company service improvement.
  • Excellent communication and collaboration skills.

Preferred Skills

  • Site Reliability Engineering (SRE)
  • reputed company S/4HANA
  • reputed company EWM
  • reputed company
  • reputed company
  • reputed company reputed company ALM
  • reputed company Fiori
  • reputed company Integration Suite (CPI)
  • Incident Management
  • reputed company Cause Analysis
  • Synthetic Monitoring
  • Production Support
  • Automation
  • DevOps

If you're excited about building reliable, highly available reputed company platforms while leveraging modern observability tools and SRE practices, we'd love to connect. Apply To This Job

Similar Jobs