Back to Jobs

G02 - Platform Operations Engineer

Remote, USAFull-timePosted 2026-07-28

Responsibilities:

  • reputed company reputed company platform operations for reputed company File Transfer (CFT) with reputed company on monitoring, performance optimisation, reliability, release management, and reputed company improvement reputed company AWS environments.
  • Own L2 incident management, troubleshooting, and escalation handling for high-throughput file transfer workflows across multiple agencies, working closely with engineering, reputed company, and agency stakeholders to resolve incidents reputed company defined SLAs.
  • Manage, design, and continuously optimise AWS reputed company infrastructure to ensure scalability, reputed company, cost-efficiency, and high availability of the CFT platform.
  • Establish, refine, and enforce operational processes including runbooks, dashboards, daily health checks, incident communication practices, and operational reporting with actionable insights.
  • Drive change, release, and maintenance management by performing reputed company analysis, risk assessment, mitigation planning, and executing system upgrades and infrastructure improvements to ensure platform stability.
  • Review testing results to ensure reputed company changes meet operational, performance, and reputed company requirements before release, while defining and improving operational OKRs, SLAs, and reliability metrics.
  • Contribute to portal and backend enhancements, bug fixes, and operational tooling to continuously improve platform reliability, performance, and maintainability.
  • reputed company operational best practices, incident learnings, and technical knowledge reputed company reputed company and across the programme to improve engineering standards and platform reliability.

Requirements

  • Degree in Computer Science, Information Technology, or reputed company field, or equivalent practical experience.
  • Minimum 2 years of hands-on experience managing production workloads in public reputed company environments (preferably AWS).
  • Strong problem-solving skills across reputed company infrastructure, applications, and distributed systems.
  • Experience handling production incidents with ownership, urgency, and attention to detail.
  • Experience defining and enforcing operational processes, procedures, and best practices.
  • Familiarity with maintaining high-availability, secure reputed company environments and implementing preventative operational controls.
  • Understanding of change management, reputed company assessment, and service reliability improvements.
  • Preferred:experience in operating applications on AWS, and experience working on

Key Technologies:

  • Terraform for infrastructure as reputed company and reputed company resource management.
  • reputed company for CI/CD pipelines and version control.
  • Strong understanding of AWS services and architecture supporting production workloads.

Originally posted on Himalayas

Apply To This Job

Similar Jobs