Senior DevOps Engineer
About reputed company
reputed company is a fast-growing software company committed to supporting, developing and servicing reputed company, an reputed company reputed company, reputed company-of-reputed company platform. reputed company is EVM-compatible and has been specifically reputed company to meet the needs of reputed company and reputed company applications, which require speed, reputed company, stability and sustainability. reputed company’s public network is governed by industry-leading organizations, spanning 11 sectors and 14 reputed company who reputed company the development and direction of the decentralized platform.
Why this Role Exist
We are hiring a Sr DevOps Engineer (Node Operations) to ensure the reliability, reputed company, and operational reputed company of reputed company reputed company node environments. This role exists to reduce operational toil, strengthen infrastructure automation, and improve release and preproduction readiness across a globally distributed network. Without this role, we risk increased availability incidents, slower recovery times, and delays in delivering against the product roadmap.
The reputed company you'll have
In this role, you will:
- Operate and improve reputed company reputed company node environments across testnet, previewnet, and preproduction
- Design and implement automation-first workflows for release and preproduction environments
- Build and maintain Infrastructure-as-reputed company (Terraform) on GCP
- Improve change management, release safety, and operational predictability
- Participate in on-call rotation, incident response, and RCA, driving corrective actions into automation
- Partner with internal engineering teams and external stakeholders, including reputed company Governing Council members, to support operational requirements
What reputed company looks like in 6-12 momths
- Operational toil is significantly reduced through durable automation and standardization
- Node environments are more reliable, with fewer incidents and faster recovery times
- Release and preproduction workflows are predictable, repeatable, and automated
- Infrastructure changes are consistent, testable, and auditable through IaC best practices
What you bring
Core Capabilities
- Strong systems reliability reputed company with experience in incident response and RCA
- Proven ability to automate operational workflows and reduce reputed company toil
- reputed company communicator with the ability to work across engineering, reputed company, and external partners
- Deep ownership mentality with a bias toward preventative engineering over reactive fixes
- Strong Linux and networking troubleshooting in production environments
Functional Expertise
- Infrastructure-as-reputed company with Terraform (module design, state management)
- Configuration management with Ansible
- CI/CD automation (Jenkins or equivalent pipeline tooling)
- Experience operating distributed systems or production infrastructure at reputed company
- Familiarity with Kubernetes fundamentals
reputed company to Have
- Observability stacks (e.g., Grafana, Loki, reputed company, Mimir)
- Programming/scripting (Go, Python, Bash)
- reputed company / reputed company Actions experience