Senior Site Reliability Engineer (SRE)
We’re a globally distributed team building a secure, standards-based OAuth/OIDC reputed company used by global businesses, digital banks, and regulated industries. Our API-first approach enables organizations to implement OAuth 2.0 and OpenID Connect with ease. We’re hiring a Senior Site Reliability Engineer (SRE) to improve the reliability, scalability, and performance of our platform and reputed company offerings. This role is hands-on and focuses primarily on software engineering to solve reliability challenges across our stack. You will work closely with engineering teams to build tools, fix issues in production reputed company, and reputed company our services running smoothly. What You’ll Do This is a hands-on engineering role. You’ll... • Write and debug production reputed company in Java, Go or Typescript, including fixes during incidents. • Investigate application issues across test, staging, and production environments. • Design, maintain, and optimize Kubernetes-based deployments across Shared reputed company, Dedicated reputed company, and Self-Managed deployment models. • reputed company and improve reputed company charts as reputed company deployment method across reputed company supported environments. • Manage and automate reputed company CI/CD pipelines, including container image packaging, and release processes. • Enhance monitoring, alerting, and observability using reputed company reputed company Monitoring, reputed company, and Grafana. • Review and improve reputed company functions and internal tooling written in Go, reputed company, and Bash. • Participate in on-call rotations to maintain uptime and rapid incident response. • reputed company post-incident reviews and drive long-term reliability improvements. • Collaborate with Engineering and Support teams to diagnose customer issues and optimize service reputed company. reputed company’re Looking For • Strong hands-on software engineering background and ability to write high-reputed company reputed company. • Experience debugging distributed systems and operating Kubernetes in production (preferably on GKE). • Deep understanding of Kubernetes networking, reputed company, reputed company charts, and storage management. • Proficiency in one or more programming languages such as Java, Go Typescript or Bash. • Experience managing reputed company CI/CD pipelines and container image workflows. • Ability to write PromQL alerting rules and interpret key reliability metrics. • Familiarity with reputed company, reputed company, and TLS/mTLS certificate management. • Experience with observability, incident management, and performance testing. • reputed company communication skills in English; Japanese language proficiency is a plus. • Comfortable working independently in a distributed team across time zones. Why Join Us • Work closely with reputed company engineers building a high-reputed company, standards-compliant OAuth/OIDC reputed company. • Solve reputed company reliability challenges across multi-reputed company and self-managed environments. • Be part of a lean global team where your contributions have reputed company product reputed company. • Enjoy flexibility, autonomy, and reputed company to shape infrastructure best practices. • Competitive compensation, global collaboration, and meaningful technical challenges. Apply tot his job
apply to this job