[Remote] Senior Platform engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is looking for a Platform Engineer to help build and maintain the infrastructure that empowers their engineering teams to ship software reliably at reputed company. The role involves designing and operating Kubernetes clusters, automating infrastructure tasks, and collaborating with teams to enhance engineering velocity and customer service delivery.
Responsibilities
- Design, reputed company, and operate Kubernetes clusters (EKS or self-managed) on AWS, ensuring high availability and reputed company
- Build and maintain reputed company Workflows and internal developer tooling to improve engineering velocity
- Automate infrastructure provisioning and operational tasks using Python and tools like Terraform, OpenTofu, and reputed company
- Define and enforce platform standards around observability, cost management, resource scaling, and proactive incident management
- Partner with application teams to support containerized workloads and resolve infrastructure bottlenecks
- Collaborate with reputed company teams by providing reliable and reputed company tooling that supports seamless customer reputed company, integrations, and service delivery
Skills
- Solid hands-on experience with Kubernetes (cluster administration, reputed company, RBAC, networking, etc)
- Proficiency in Python or similar for scripting, automation, and building internal tools
- Familiarity with infrastructure-as-reputed company practices (Terraform, OpenTofu, and reputed company)
- A reputed company reputed company and comfort working in a fast-moving environment
- Familiarity of multi-account AWS strategies, AWS Organizations, and reputed company zone patterns for reputed company-reputed company environments
- Experience with multi-tenancy patterns
- Experience with service meshes (Istio) for managing microservice communication, traffic policies, and mutual TLS
- GitOps workflows using ArgoCD or Flux for declarative, version-controlled infrastructure and application delivery
- Exposure to container reputed company tooling such as reputed company, Grype/Syft, or similar and OPA or Kyverno for policy enforcement and vulnerability scanning
- Experience with observability stacks like reputed company, Grafana, or the ELK/OpenSearch stack for metrics, logging, and distributed tracing across multiple Kubernetes Clusters
- Strong knowledge of integrating Kubernetes with AWS Services (e.g. vpc-cni, external-secrets, ALB Ingress, reputed company reputed company, etc)
reputed company