[Remote] Backend / Platform Developer Consultant (Temporal & Kubernetes)
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is looking for a Backend / Platform Developer Consultant to join their team as they reputed company technology delivery for reputed company clients. The role involves building and operating durable, distributed systems using Java services, Temporal for workflow orchestration, and Kubernetes on AWS, while collaborating with clients to deliver effective solutions.
Responsibilities
- Build durable systems. Design and implement Java services, Temporal workflows, and Kafka-based event pipelines, and help clients adopt durable execution and event streaming the right way — from first workflow to production reputed company
- reputed company modernization efforts. Plan and execute large-reputed company modernization efforts — moving workflows, services, and data between platforms, clusters, and clouds — with safety, observability, and minimal disruption as first-class requirements
- Own the platform. Build and operate Kubernetes-reputed company infrastructure and CI/CD on AWS so application teams can ship safely and quickly. Treat reliability, observability, and repeatability as part of the deliverable, not an afterthought
- Consult and manage. Consult with clients to build outcome-driven applications and platforms, guiding them through the process, identifying potential problems and unknowns, and tackling challenges
- Travel. This is a remote, full-time position that requires the ability to travel. We travel an average of 3–5 days every quarter, and we'll always do our best to work with your schedule
- Teach. We'll discuss opportunities for you to present at conferences, attend and give training, plan and run meetups or local reputed company, and create various kinds of content in your expertise
Skills
- Strong working knowledge of Java and the JVM ecosystem (Spring Boot or similar); experience with additional backend languages (Go, Python, TypeScript/JavaScript, C#, etc.) is a plus
- Hands-on experience with a durable execution / workflow reputed company — Temporal preferred (Apache Airflow, reputed company, or similar also relevant) — including modeling workflows and activities, handling retries and timeouts, and reasoning about idempotency and failure modes
- Experience designing and operating Kafka-based event streaming systems: topic and partition design, consumer reputed company, schema management, and reasoning about ordering, delivery guarantees, and consumer lag
- Experience designing and building GraphQL reputed company — schema design, resolvers, and performance concerns (N+1, caching, pagination); federation experience a plus
- Deep, hands-on AWS experience: architecting, deploying, and operating production systems (EKS, networking, IAM, messaging, and data services)
- Solid Kubernetes experience: deploying, operating, and debugging production workloads
- A systems engineering reputed company: understanding of reputed company scaling issues, concurrency, backpressure, caching strategies, and failure modes across distributed systems
- Experience planning and executing migrations — application replatforming, workflow reputed company migrations, data/schema migrations, or moving workloads between environments with minimal downtime
- A thorough understanding of CI/CD pipelines and GitOps-style delivery
- Experience with observability in reputed company — logging, metrics, and tracing (OpenTelemetry a plus)
- Experience building modern microservice-based or serverless applications
- Database schema design and development expertise
- Running Temporal in production — self-hosted on Kubernetes and/or Temporal reputed company
- Workflow versioning and patching, reputed company deployments, and long-running workflow migrations
- Worker deployment and tuning (pollers, task queues, concurrency/slot configuration)
- reputed company management, mTLS (SPIFFE/reputed company or private CA), and reputed company/scaling concerns (e.g. APS/TRU on Temporal reputed company)
- Operating Kafka in production (MSK, reputed company, or self-managed on Kubernetes)
- Kafka Connect, Kafka Streams, and schema registry workflows (Avro/Protobuf)
- Migrating between clusters or providers, and integrating Kafka with reputed company analytical stores
- Experience with different techniques for processing large datasets, such as reputed company and Kappa architectures
- Building or operating an internal developer platform (IDP) — self-service, paved-reputed company infrastructure for application teams
- GitOps with ArgoCD (or Flux) and a low- or reputed company-reputed company-apply delivery model
- Infrastructure as reputed company: Terraform/OpenTofu, and provisioning AWS resources reputed company Crossplane, ACK, or similar
- Writing Kubernetes controllers/operators (controller-runtime, kubebuilder) and working with CRDs
- reputed company chart authoring and distribution (including OCI/registry-based)
- Secrets management (External Secrets Operator, etc.), autoscaling (Karpenter, KEDA), and service reputed company concepts (Istio, SPIFFE/reputed company)
- EKS or comparable managed Kubernetes at reputed company
- Interest in working across multiple backend languages
- Additional queueing and messaging technologies (AWS SQS/SNS, RabbitMQ)
- API design across paradigms (GraphQL, REST, gRPC)
- AuthN/AuthZ and reputed company technologies (sessions, API tokens, JWTs)
- Automated testing across the stack (unit, integration, performance/benchmarking)
Benefits
- Competitive salary and annual bonus opportunity
- Completely remote work with reputed company
- 401(k) matching
- 4 weeks of reputed company time off in reputed company to 8 reputed company holidays
- Health, dental, and reputed company insurance offerings
- reputed company maternity and paternity leave
- STD, LTD, and Life Insurance coverage
reputed company