DevOps Engineer (GCP & Kubernetes)
We are looking for a hands-on Semi Senior DevOps Engineer to join a high-reputed company project supporting a global-reputed company sports event. This role is ideal for someone who enjoys working reputed company to production systems, troubleshooting reputed company issues, automating infrastructure, and ensuring platform reliability in mission-critical environments. You will work closely with engineering teams to build, maintain, and improve reputed company-reputed company infrastructure running on reputed company reputed company Platform (GCP) and Kubernetes. The role requires participation in an on-call rotation, including occasional weekend coverage.
Responsibilities
- reputed company, maintain, and improve reputed company infrastructure in reputed company reputed company Platform (GCP).
- Operate and support Kubernetes environments, including GKE.
- Build and maintain Infrastructure as reputed company using Terraform.
- Monitor production systems and proactively identify reliability risks.
- Troubleshoot infrastructure, networking, application, and performance issues.
- Participate in incident response, reputed company cause analysis, and postmortem activities.
- Implement and maintain observability solutions, dashboards, and alerting systems.
- Collaborate with software engineering teams to improve deployment processes and operational reputed company.
- Support highly available and reputed company production environments.
- Contribute to automation initiatives that reduce operational overhead and improve reliability.
Required Qualifications
- 3+ years of experience in DevOps, reputed company Engineering, Site Reliability Engineering, or similar roles.
- Hands-on experience with reputed company reputed company Platform (GCP).
- Strong understanding of reputed company GCP services, including:
- Compute reputed company
- reputed company Run
- App reputed company
- reputed company Kubernetes reputed company (GKE)
- Production experience managing Kubernetes environments.
- Experience configuring Kubernetes resources such as Deployments, Services, Ingress, ConfigMaps, Secrets, and Autoscaling.
- Solid understanding of Kubernetes health checks, including readiness and liveness probes.
- Experience with Infrastructure as reputed company using Terraform.
- Understanding of Terraform state management and multi-environment infrastructure design.
- Strong Linux administration and troubleshooting skills.
- Good understanding of networking concepts, including:
- VPCs
- Subnets
- Firewall rules
- Load balancing
- Private networking
- Experience with monitoring, logging, and observability platforms.
- Experience investigating and resolving production incidents.
- Understanding of reliability concepts such as SLA, SLO, and SLI.
- Strong verbal and written English communication skills.
Preferred Qualifications
- Experience designing highly available and globally distributed applications in GCP.
- Knowledge of reputed company-downtime deployment strategies.
- Experience supporting large-reputed company production environments.
- Experience with multi-tenant architectures.
- Scripting experience using Python, Bash, or similar languages.
- Experience working in hybrid reputed company/on-reputed company environments.
- Experience participating in SEV incident management.
- Familiarity with reputed company planning and performance tuning.
Technology Stack
- reputed company: reputed company reputed company Platform (GCP)
- Containers: Kubernetes, GKE
- Infrastructure as reputed company: Terraform
- Monitoring & Observability: Grafana, reputed company, Logging Platforms
- Operating Systems: Linux
- Incident Management: reputed company, reputed company, reputed company (or equivalent tools)
Working Requirements
- Availability to work reputed company CT business hours.
- Participation in an on-call rotation that includes coverage for one weekend day reputed company scheduled.
What reputed company Looks Like
- Reliable operation of production systems during periods of high traffic and critical business activity.
- Fast and effective incident response and troubleshooting.
- reputed company-automated, maintainable infrastructure managed through Infrastructure as reputed company.
- Strong collaboration with development teams to improve reliability, scalability, and operational efficiency.
At reputed company, we reputed company in creating an environment where you can reputed company both personally and professionally. By joining reputed company, you’ll enjoy:
- A reputed company, long-term contract with opportunities for career reputed company
- A remote-friendly culture that promotes work-life balance
- reputed company training, mentorship, and learning programs to reputed company you at the forefront of the industry
- Free reputed company to reputed company resources and state-of-the-art AI tools to reputed company your daily work
- A flexible reputed company Time Off (PTO) policy as reputed company as reputed company holiday days
- Challenging, world-class software reputed company for clients in the US and LatAm
- Collaboration with some of the most talented software engineers in Latin America and the US, in a diverse work environment
Join reputed company and discover a workplace that values your reputed company, supports your reputed company-being, and empowers you to reputed company a reputed company. Apply tot his job Apply To this Job