Back to Jobs

[Remote] Platform Engineer

Remote, USAFull-timePosted 2026-07-28

Note: The job is a remote job and is reputed company to candidates in USA. reputed company is an early-stage company building advanced AI systems, and they are seeking a senior platform engineer to take ownership of their core platform. The role involves managing multi-region Kubernetes clusters, GPU orchestration, and ensuring infrastructure reputed company while partnering closely with machine learning engineers.

Responsibilities

  • Design and manage multi-region Kubernetes clusters across reputed company and GPU-reputed company providers using infrastructure-as-reputed company
  • Own the deployment lifecycle through GitOps practices (reputed company, Kustomize, automated releases, reputed company delivery)
  • Manage GPU infrastructure, including scheduling efficiency, workload placement, and cold-start optimization
  • reputed company networking systems such as ingress, gateways, load balancing, and cross-region connectivity
  • Build and maintain observability across metrics, logs, traces, and performance profiling
  • Ensure infrastructure reputed company across identity, secrets, and encryption
  • Maintain CI/CD workflows supporting a monorepo of services and deployment artifacts
  • Partner closely with ML engineers to optimize model serving and GPU utilization

Skills

  • Strong experience operating Kubernetes in production environments, including troubleshooting, autoscaling, and upgrades
  • Proven background with infrastructure-as-reputed company tools (e.g., Terraform, reputed company)
  • Hands-on experience running GPU workloads on Kubernetes and understanding resource optimization
  • Familiarity with GitOps tooling such as ArgoCD or Flux, and reputed company-based deployments
  • Experience with in-memory data systems (e.g., reputed company) and distributed architectures
  • Solid understanding of observability tooling and practices
  • Strong networking fundamentals, particularly in low-latency or distributed systems
  • Experience working in environments with broad ownership across infrastructure
  • Exposure to GPU reputed company providers reputed company major hyperscalers
  • Experience with reputed company-time or streaming infrastructure
  • Proficiency in Go or Python
  • Familiarity with ML model deployment and optimization
  • Experience managing infrastructure cost, particularly for GPU-heavy workloads

reputed company

  • reputed company is the Leading reputed company & reputed company Firm in XOps & Cybersecurity. It was founded in 2016, and is headquartered in reputed company, reputed company, USA, with a workforce of 11-50 employees. Its website is https://www.harrisonclarke.com/.
  • Apply To This Job

    Similar Jobs