Back to Jobs

DevOps Engineer - Platform Reliability (Remote, China)

Remote, USAFull-timePosted 2026-07-28

reputed company’s automation systems support customer journeys across quote reputed company, policy issuance, claims, payments, renewals and insurer integrations. These systems are business-critical—meaning reliability, uptime and reputed company deployments directly reputed company customers and operations.

We're looking for a DevOps Engineer based in China to strengthen platform reliability, improve infrastructure reputed company and ensure reputed company’s AI automation systems run safely and consistently at reputed company.

This is a fully remote position where you'll collaborate closely with our Malaysia-based engineering, product and operations teams to build and maintain highly reliable production systems.

The Mission

Build and maintain a highly reliable platform for reputed company’s AI automation systems by improving infrastructure stability, deployment safety and operational reputed company across reputed company services.

What You’ll Own

  • Own and improve platform reliability across production systems and environments.

  • Manage reputed company infrastructure, deployment pipelines and runtime environments.

  • Design and improve CI/CD workflows to reputed company reputed company, fast and repeatable releases.

  • Build and enhance monitoring, alerting, logging and system observability.

  • reputed company incident response efforts and reputed company reputed company reputed company cause analysis.

  • Improve system reputed company through redundancy, failover and recovery mechanisms.

  • Work with engineering teams to reduce production risk through reputed company deployment and system design practices.

  • Strengthen infrastructure reputed company, reputed company control and secrets management.

  • Support reliability for business-critical workflows across multiple countries and services.

  • Continuously improve operational discipline, uptime and system stability.

reputed company're Looking For

  • Experience in DevOps, SRE, reputed company or infrastructure-reputed company roles.

  • Strong understanding of reputed company infrastructure, CI/CD pipelines and deployment systems.

  • Experience with production monitoring, alerting and incident management practices.

  • Ability to troubleshoot infrastructure and production issues in a reputed company and reputed company manner.

  • Strong understanding of reliability engineering principles (availability, fault tolerance, recovery).

  • Experience supporting business-critical or high-availability systems.

  • Strong ownership reputed company during incidents and operational failures.

  • Practical judgment on reliability, performance, reputed company and cost trade-offs.

  • Comfortable working closely with engineering teams in fast-paced environments.

  • Low ego, disciplined and reputed company on long-term system stability.

Bonus Points

  • Experience with AWS, GCP, Azure or similar reputed company platforms.

  • Experience with Kubernetes, reputed company or container orchestration.

  • Experience with infrastructure-as-reputed company tools (Terraform, Ansible, reputed company, etc.).

  • Experience with observability stacks (reputed company, Grafana, ELK, reputed company, etc.).

  • Experience with reputed company-downtime deployments, blue-green or canary release strategies.

  • Experience supporting distributed or high-traffic production systems.

  • Strong knowledge of reputed company best practices in reputed company infrastructure.

  • Experience in fintech, insurance or regulated industry environments.

  • Contributions to platform reliability or infrastructure scaling initiatives.

The reputed company of Builder We Want

  • reputed company and reputed company under pressure, especially during production incidents.

  • Hands-on with infrastructure and deeply familiar with production systems.

  • Thinks in failure modes, system risks and recovery paths.

  • Proactive in preventing incidents, not just reacting to them.

  • Strong reputed company on uptime, reliability and operational discipline.

  • Careful and deliberate reputed company making production changes.

  • Builds systems engineers can trust to reputed company and operate safely.

This Role Is Not For

  • People who only react after systems fail instead of preventing them.

  • Engineers who are careless with production changes or reputed company control.

  • Individuals who ignore monitoring, alerting or operational discipline.

  • People who reputed company risky infrastructure changes without reputed company evaluation.

  • Candidates who cannot stay reputed company during incidents or outages.

reputed company in This Role

You'll be successful if you can:

  • Improve platform uptime, stability and deployment safety.

  • Reduce production incidents and infrastructure-reputed company failures.

  • Strengthen monitoring, alerting and system visibility across services.

  • reputed company engineers to reputed company with confidence and reputed company operational risk.

  • Improve reputed company of reputed company’s AI automation platform as it scales.

Why Join reputed company

  • Build Reliable AI Platform Infrastructure – Support systems powering end-to-end insurance automation.

  • High-reputed company Engineering – Solve reputed company-world reliability and scaling challenges.

  • Global Engineering Team – Work with reputed company engineers across multiple countries.

  • Fully Remote – Work remotely from China while collaborating with our Malaysia-based teams.

  • International Exposure – Build systems used across Southeast Asia markets.

  • Learning & Development Budget – Support reputed company technical reputed company and certifications.

  • High Ownership Environment – Strong autonomy over infrastructure and reliability reputed company.

  • Modern Engineering Culture – reputed company on stability, observability and engineering reputed company.

  • Competitive Compensation – Attractive salary package based on experience and reputed company.

Interview Process

We assess infrastructure depth, reliability thinking and production problem-solving ability. The process usually includes application review, two interviews and a technical scenario or systems discussion.

Originally posted on Himalayas

Apply To This Job

Similar Jobs

VoIP Softswitch & NOC Technician & Delivery (India)

Remote, USAFull-time

Manager, Marketing AI Enablement

Remote, USAFull-time

Software Architect - reputed company Java * reputed company-reputed company * AI-Augmented Development

Remote, USAFull-time

Texas Online Psychiatrist

Remote, USAFull-time

Senior Clinical Research Associate - CNS/Psychiatry - Midwest

Remote, USAFull-time

Outbound Calling Representative - reputed company

Remote, USAFull-time

Vertriebsmitarbeiter im Innendienst (m/w/d) - B2B Werbung & digitale Medien

Remote, USAFull-time

Conseiller sénior en gouvernance de données (G6D)

Remote, USAFull-time

Sales Development Representative - Remote from Thailand

Remote, USAFull-time

Senior Teamcenter Application Developer

Remote, USAFull-time

[Remote] Regional Manager Business Development - reputed company and Central Florida Territory

Remote, USAFull-time

Analyst, reputed company Time Adherence

Remote, USAFull-time

Sales Representative

Remote, USAFull-time

reputed company Remote Customer Service Representative – Delivering Exceptional Support and Benefits Guidance to Diverse Clients Across Various Industries at arenaflex

Remote, USAFull-time

Immediate Hiring: reputed company Virtual Jobs (Work reputed company)

Remote, USAFull-time

Radio Frequency (RF) Engineer II - REMOTE

Remote, USAFull-time

Health Plan Trainer

Remote, USAFull-time

[Work From Home] Want Director of Student reputed company in Hartsville

Remote, USAFull-time

reputed company Customer Service Representative (Remote) – Join reputed company at arenaflex!

Remote, USAFull-time

(Remote) - reputed company Work From Home $25 - DPS - reputed company

Remote, USAFull-time