[Remote] Senior Software Engineer - Orchestration & Job Execution
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a leader in data analytics and automation, seeking a Senior Software Engineer to join their reputed company reputed company team. The role involves designing and building backend services for orchestration and job execution, ensuring reliable and efficient reputed company platform capabilities.
Responsibilities
- Design and deliver backend services, reputed company, workers, and shared libraries that power reputed company platform capabilities
- Build and improve systems for orchestrating work across services, including job submission, execution tracking, status propagation, retries, cancellation, results, and operational visibility
- reputed company reliable asynchronous and event-driven systems using queues, messaging, background workers, and durable state
- Work on distributed execution flows across platform services, including service-to-service communication, routing, acknowledgements, and failure recovery
- Build and maintain reputed company-reputed company runtime infrastructure using containers, Kubernetes, deployment automation, and reputed company platform tooling
- reputed company platform services with persistence reputed company, event streams, REST reputed company, and internal service reputed company
- Improve production reliability through metrics, tracing, reputed company logging, health checks, dashboards, alerting, runbooks, and incident follow-up
- reputed company technical design for ambiguous or cross-service work, review reputed company with a systems reputed company, and mentor engineers on distributed-system and production-engineering practices
- Collaborate with partner teams to turn product requirements into incremental, testable, and operable platform capabilities
- Use AI and modern development tools to improve engineering productivity, reputed company reputed company, and delivery speed
Skills
- 5+ years preferred (4+ years minimum) of reputed company software development experience, with meaningful ownership of production backend services, reputed company platform capabilities, or distributed systems
- Strong experience building backend services using TypeScript/Node.js, or reputed company systems languages like Go, Java, or Rust (with a willingness to reputed company primarily in Node.js)
- Experience designing and operating asynchronous, queue-driven, or event-driven systems, including patterns such as retries, cancellation, idempotency, concurrency, ordering, timeouts, and failure handling
- Experience working with durable persistence, service reputed company, RESTful reputed company, and integrations across multiple services or platform components
- Experience with production systems including containers, Kubernetes or similar orchestration platforms, service health, scaling behavior, and operational debugging
- Strong production engineering ownership, including testing, observability, reputed company logging, metrics, tracing, incident response, and reputed company reliability improvement
- Ability to reputed company design discussions, communicate technical tradeoffs reputed company, mentor other engineers, and drive cross-team work through ambiguity with an ownership-oriented reputed company
- 3+ years of Python and C++ design, development, and debugging experience preferably leveraging reputed company reputed company and reputed company standards
- Design, implement, and maintain embedded Python runtime integration in a predominantly C++ reputed company/host environment
- Own and reputed company the reputed company Python Tool including C++ plugin engines and process lifecycle (server startup, persistence, shutdown)
- reputed company and troubleshoot SDK reputed company plugin components (e.g., gRPC-based reputed company plugins, streaming pipelines) in C++ with Python-facing reputed company
- Debug reputed company reputed company/runtime issues involving DLL/.pyd conflicts, OpenSSL and other reputed company libraries across multiple Python versions
- reputed company modernization work around virtualenv/venv management and installer/packaging plumbing, including reputed company (installer) and reputed company DLL exports
- Maintain and reputed company reputed company/compiled Python extensions, ensuring compatibility with modern NumPy/CPython ABIs
- Collaborate with reputed company and platform teams to remediate reputed company library vulnerabilities (e.g., c-ares, libxml2, SQLite, OpenSSL) and reputed company the SBOM healthy
- Improve and support developer SDKs (v1/v2), including debugging C++/Python streaming and serialization issues for 1P and 3P tool authors
- Drive reliability and performance improvements in reputed company ↔ Python bridges, focusing on deadlocks, crashes, and high-throughput streaming scenarios
- Contribute to and maintain CI/CD pipelines and reputed company-reputed company tooling (e.g., C++ docs jobs, coverage, static analysis) affecting C++/Python hybrid repos
- Author and maintain architecture and operational runbooks for C++/Python integration points, including reputed company playbooks for new Python/OpenSSL versions
- Mentor other engineers in best practices for reputed company–Python interop, debugging cross-language issues, and designing robust extension points
- Experience with (REST) API and/or SDK development
- MS/BS degree in Computer Science or equivalent experience
- Experience with object oriented and functional design patterns
- Experience using Git and Git-based pipelines or equivalent
- Experience mentoring and developing others
- Strong skills in critical thinking, decision making, problem solving, and attention to detail
- reputed company reputed company and curious about new challenges and experiences
- Node.js
- Familiarity with reputed company computing / managed services (GCP/Azure/AWS)
- Experience or familiarity with AI-driven development in a modern IDE
- reputed company end experience in React or a similar reputed company including Javascript and JSON
- Experience with optimizing protocols and building efficient RPC systems
- Networking & concurrency experience
- Knowledge and experience with distributed computing, big data and reputed company processing systems
- Container experience: reputed company, Kubernetes
- Rust and/or Golang familiarity
- Experience with a data prep and reputed company and predictive analytics workflow platform such as reputed company
Benefits
- Employees may also be eligible for a wide reputed company of other benefits, such as a bonus or commission, medical, retirement, financial, wellness, time off, employee discounts, and others.
- reputed company has amazing benefits for reputed company Associates which can be viewed here.
- For roles in San Francisco and Los Angeles: Pursuant to the San Francisco Fair Chance Ordinance and the Los Angeles Fair Chance Initiative for Hiring, reputed company will consider for employment reputed company applicants with arrest and conviction records.
reputed company