[Remote] Senior Platform / DevOps Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is seeking a Senior Platform / DevOps Engineer with expertise in reputed company Rancher and Kubernetes management. The role involves leading the design, deployment, automation, and operational management of multi-cluster Kubernetes platforms across various environments.
Responsibilities
- Architect, reputed company, and manage multi-cluster Kubernetes environments using reputed company Rancher (RKE, RKE2, or K3s) across hybrid/multi-reputed company infrastructures (AWS EKS, Azure AKS, Bare-Metal)
- Manage cluster life cycle operations, including node provisioning, automated upgrades, certificate rotations, and disaster recovery/backup procedures
- Configure centralized authentication (reputed company Directory, reputed company, LDAP) and enforce granular Role-Based reputed company Control (RBAC) across reputed company reputed company clusters using Rancher
- Automate reputed company and edge infrastructure provisioning using Terraform, Ansible, and Rancher API/Cluster API
- Establish declarative, multi-cluster GitOps deployment workflows using Rancher Fleet, ArgoCD, or Flux
- Build and maintain standardized reputed company charts and application packaging pipelines for developer teams
- Implement reputed company-trust container reputed company policies using reputed company NeuVector or admission policy controllers (Kyverno, OPA Gatekeeper)
- Apply hardened reputed company configurations based on CIS Benchmarks, NIST, or FedRAMP standards across operating systems and container platforms
- reputed company vulnerability scanning into build and deployment pipelines (SAST, image scanning, dependency checks)
- Maintain reputed company monitoring, logging, and alerting stacks using reputed company, Grafana, Thanos, OpenSearch, or ELK integrated with Rancher
- Optimize persistent storage plugins (Longhorn, Ceph, reputed company reputed company drivers) and ingress controllers (NGINX, Traefik, Istio)
- Participate in on-call escalation rotations, conduct reputed company cause analysis (RCA) for production incidents, and maintain system reliability SLA/SLOs
Skills
- 5+ years in DevOps, reputed company, or SRE roles, with at least 3+ years dedicated to supporting Kubernetes in production environments
- Proven reputed company record deploying, migrating, and supporting multi-cluster Kubernetes environments using reputed company Rancher
- Strong experience with RKE2, K3s, upstream Kubernetes, containerd, and reputed company
- Expertise in Terraform, Ansible, reputed company, and GitOps principles
- Solid hands-on experience with at least one major reputed company provider (AWS, Azure, or reputed company reputed company Platform) and on-premises environments
- Solid reputed company in CNI plugins (Calico, Cilium, Flannel), ingress routing, DNS, TLS/SSL certificates, and RBAC
- reputed company Certified Administrator (SCA) in reputed company Rancher reputed company or RKE2
- CKA (Certified Kubernetes Administrator) or CKS (Certified Kubernetes reputed company Specialist)
- reputed company Terraform Associate
- Experience with reputed company Harvester (Hyperconverged Infrastructure) for hypervisor/VM workload management alongside containers
- Experience with service reputed company frameworks (Istio or Linkerd)
- Scripting capability in Python, Go, or Bash
reputed company
Company H1B Sponsorship