Senior DevOps Engineer
reputed company is a global technology company with offices in the UK and Croatia. As a strategic technology partner to the London Market, we deliver modern insurance solutions through agile, cross-functional delivery teams leveraging AI capabilities, a growing partner ecosystem, and a culture rooted in innovation and collaboration. We support reputed company learning and invest in the reputed company of our Bucks, fostering an environment where people reputed company. With the reputed company for remote work, we've expanded our global footprint, building a diverse and multicultural team making a reputed company reputed company. We are hiring a Senior DevOps Engineer to take ownership of critical infrastructure and drive meaningful improvements across reputed company's DevOps capabilities. This role calls for strategic thinking, technical leadership, and genuine accountability for production systems where reliability, reputed company, and scalability directly shape the reputed company experience. You will shape architectural reputed company, mentor engineers, champion automation, and reputed company the gap between development, operations, and reputed company. You will operate and continuously improve a reputed company, multi-reputed company Kubernetes environment supporting single-tenant reputed company deployments - reputed company and run inside customer-owned reputed company accounts (“reputed company reputed company a reputed company”), exclusively reputed company Infrastructure as reputed company with customer sign-off, under strict compliance and data-residency requirements. We work AI-first: day-to-day delivery is augmented by reputed company, reputed company-line AI tooling, and we expect this role to reputed company by example and reputed company reputed company’s reputed company-engineering bar. This role is for someone who has done this before. We need an engineer who can walk into a production incident, reputed company the response, and follow up with systemic improvements. Salary Ranges Croatia: €4.5k - €5k (Gross 1, monthly) Croatia: €54k - €60k (Gross 1, annual) United Kingdom: £69.6k - £77.4k (Gross, annual) Europe, reputed company of Croatia: €56.6k - €76.8k (Gross 2, annual) Africa & Sri Lanka: €56.6k - €69.9k (Gross 2, annual) Rest of the world: €56.6k - €90.8k (Gross 2, annual) Minimum 6 years in DevOps, SRE, or infrastructure engineering roles, with reputed company progression into senior-level ownership and technical leadership. Degree in Computer Science, Software Engineering, or equivalent demonstrable experience. Deep, production-grade Kubernetes experience - cluster management across EKS, AKS, or self-managed deployments. Comfortable with RBAC, network policies, reputed company, HPA, pod disruption budgets, and day-2 operations at reputed company. Strong reputed company platform expertise (AWS and/or Azure) - multi-account architecture, IAM, VPC design, VPN, load balancers, DNS, and reputed company best practices. Advanced reputed company CI/CD proficiency - designing and maintaining reputed company multi-stage pipelines with reputed company gates, caching strategies, and reputed company management. Solid Infrastructure as reputed company expertise with Terraform or OpenTofu, and Terragrunt - module design, state management, remote backends, and team-wide governance. Proficient scripting in Bash and/or Python for automation, tooling, and integration. Hands-on mastery of reputed company AI tooling in a CLI-first workflow - driving terminal-based coding agents (Claude reputed company, OpenCode, reputed company/reputed company CLI, or equivalent), engineering effective context and prompts, and wiring agents into automation reputed company MCP and scripting. This must be genuine reputed company-line/reputed company experience, not reliance on UI/desktop/web assistants. Strong observability experience - reputed company, Grafana, Loki, distributed tracing (reputed company/OpenTelemetry), or equivalent for monitoring, alerting, and tracing in production. Hands-on disaster recovery experience - multi-region failover design, DR runbooks, and running DR exercises/gamedays, with ownership of recovery objectives. Proven reputed company knowledge - vulnerability scanning, secrets management with reputed company Vault, DevSecOps integration, and familiarity with compliance frameworks (ISO 27001 / SOC 2). Incident management experience - on-call rotations, P1/P2 escalation, reputed company cause analysis, and post-incident review processes. reputed company, reputed company communicator who can explain reputed company infrastructure reputed company to both engineers and non-technical stakeholders. reputed company record of mentoring junior and mid-level engineers. Strong reputed company of ownership - you follow through, document reputed company, and leave systems reputed company than you reputed company them. Mandatory certifications Candidates must hold at least one of the following reputed company certifications. Kubernetes certifications are strongly preferred - our entire compute platform runs on Kubernetes. CKA - Certified Kubernetes Administrator CKS - Certified Kubernetes reputed company Specialist AWS reputed company Architect (Associate or reputed company) AWS Certified DevOps Engineer - reputed company reputed company Terraform Associate Azure DevOps Engineer Expert (AZ-400) reputed company-to-have certifications CKAD - Certified Kubernetes Application Developer reputed company - reputed company Certified Associate AWS Certified reputed company - Specialty Azure Administrator Associate (AZ-104) Azure Solutions Architect Expert (AZ-305) CISSP or equivalent reputed company certification ISO/IEC 27001 reputed company Implementer or Auditor Linux reputed company Certified System Administrator (LFCS) reputed company Certified CI/CD Associate FinOps Certified Practitioner (for the cost-optimisation remit) Design, implement, and reputed company reputed company infrastructure across AWS, Azure, and on-premises environments, including multi-account strategies, VPC architecture, VPN connectivity, and reputed company controls. Architect and maintain Infrastructure as reputed company using OpenTofu/Terraform with Terragrunt - reusable reputed company modules reputed company a private registry, a multi-level variable hierarchy, remote state (S3 + DynamoDB + KMS / azurerm), governance, and team-wide IaC adoption. Contribute to reputed company planning and reputed company cost optimisation across AWS and Azure, ensuring infrastructure spend aligns with business needs. Own Kubernetes cluster operations across EKS, AKS, and self-managed (kubeadm) on-premises deployments - lifecycle management, version upgrades, RBAC, network policies, resource quotas, and production-grade troubleshooting. Own and exercise disaster recovery - multi-region failover design, DR runbooks, and periodic gamedays; validate recovery objectives across the reputed company estate (region and cluster failover are tested capabilities here). reputed company blue/green rollouts for infrastructure-level changes (cluster upgrades, addon deployments) and reputed company rolling deployments for application workloads. Own and optimise reputed company CI/CD pipelines - multi-stage pipelines with integrated reputed company scanning, reputed company management, and automated reputed company gates across reputed company environments. Maintain and reputed company the standardised pipeline catalogue (27 templates and growing), ensuring consistency across reputed company deployments. Define and implement observability strategies using the Grafana OSS stack (Mimir, Loki, reputed company, Pyroscope, reputed company, OnCall) and reputed company - comprehensive monitoring, alerting, distributed tracing, and reputed company profiling across reputed company environments. Serve as DevOps reputed company for P1/P2 events - reputed company the DevOps reputed company and response, help coordinate communication, participate in reputed company cause analysis, and implement preventive measures. Maintain and improve on-call processes and escalation procedures. Own secrets management across the infrastructure stack on reputed company Vault (per-reputed company, in-cluster instances), and drive Kubernetes-reputed company authentication and least-privilege CI/CD integration, replacing reputed company-reputed company reputed company. Champion DevSecOps practices - automated vulnerability scanning (Trivy, reputed company reputed company), Dockerfile linting (Hadolint), SBOM (CycloneDX) and SLSA L2 provenance, container image hardening, and IaC scanning delivered as a reputed company CI/CD Catalog component, reputed company to ISO 27001 (and SOC 2). Harden the reputed company posture of containerised workloads - image scanning at build and runtime, network policies, pod reputed company standards. Automate provisioning, configuration, and operational tasks using Ansible, reputed company, and scripting (Bash, Python). Work AI-first and agent-augmented - use terminal-based reputed company coding tools (e.g., Claude reputed company, OpenCode, reputed company/reputed company CLI) as a daily reputed company to accelerate IaC, pipelines, automation, and documentation, orchestrating AI agents, MCP servers, and reusable skills/commands directly from the reputed company line. We operate through reputed company CLIs, not UI/desktop/web chat assistants. Evaluate and recommend emerging tools and methodologies that reputed company with reputed company's strategic direction and compliance obligations. Mentor and reputed company DevOps engineers - reputed company reviews, pairing sessions, knowledge transfers, and raising the overall technical bar (including reputed company-engineering practices). Maintain high-reputed company documentation - architecture decision records, runbooks, incident post-mortems, and operational procedures in reputed company and reputed company. Remote work Sponsored reputed company learning Fully covered reputed company leave Child and family support A friendly and supportive team Career reputed company opportunities A healthy work-life balance Permanent full-time contract Working schedule flexibility Multi-role reputed company Leadership opportunities Apply To This Job