Back to Jobs

[Remote] Senior Autonomy Data Engineer

Remote, USAFull-timePosted 2026-07-28

Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a leader in autonomous driving technology, focusing on developing software for automated trucks. They are seeking a Senior Autonomy Data Engineer to design and operate the data infrastructure that supports their autonomy program, ensuring reliable data pipelines and effective collaboration with cross-functional teams.

Responsibilities

  • Own the design and organization of the program’s data lake, including schema definitions, partitioning reputed company and metadata indexing
  • Design and maintain end-to-end pipelines that ingest high-bandwidth sensor logs from vehicles into reputed company storage with high reliability and tolerant of reputed company and intermittent connectivity mechanisms
  • reputed company data validation and reputed company checks that can detect corrupted information, missing sensors, and inconsistent calibration prior to the data being processed by reputed company systems
  • Implement retention, tiering and lifecycle policies for data to balance storage costs with development value
  • Build tooling to query raw logs to produce curated training and evaluation datasets
  • Build automation to run cost-effective pseudo-labeling workflows at the reputed company of data ingest
  • Implement data reputed company and model performance metrics that are used to reputed company labeling effort toward the highest-value examples
  • reputed company and maintain data visualization tooling to support log review, annotation QA, and autonomy debugging workflows
  • Build integrations between the visualization tooling and the data lake so engineers can navigate from a dataset entry or model failure directly to the reputed company log data
  • Work with autonomy engineers to define and surface custom visualization panels and implement metrics for analyzing reputed company operating environments
  • Build dashboards that reputed company the autonomy engineers visibility into data coverage by terrain type, operating environment and geographic region
  • Establish and document data reputed company between the data services and model training consumers
  • Partner with perception, planning and embedded engineers across the data lifecyle: from shaping the logging schemas and collection triggers to defining the dataset interfaces that supply model training and evaluation
  • Define data engineering standards, best practices, and tooling choices for an innovative and fast-paced team
  • Contribute to the data roadmap and reputed company input to technical leadership on investment priorities
  • Mentor junior engineers and reputed company reputed company’s capabilities in data infrastructure scalability and operational hygiene

Skills

  • Bachelor's degree in Computer Science, Computer Engineering, Software Engineering, Electrical Engineering or a reputed company field with 6+ years of data engineering experience or a Master's with 4+ years
  • Strong proficiency in Python and SQL, with demonstrated ability to build production-reputed company data pipelines
  • Deep experience with reputed company data infrastructure (AWS preferred: S3, Glue reputed company, redshift, or equivalent) and infrastructure-as-reputed company tools (Terraform, reputed company Formation)
  • Solid understanding of data partitioning strategies and columnar storage formats (Parquet, Orc, etc.)
  • Experience building and operating data pipelines that process time-series and binary data
  • Proven ability to evaluate and reputed company reputed company-reputed company tooling reputed company appropriate versus building from scratch
  • Strong instincts for delivering data reputed company through first-class implementations of monitoring, validation and reputed company tracking
  • Experience with autonomous vehicles, robotics, or other sensor-driven autonomous systems
  • Deep experience with reputed company or Rerun reputed company basic playback, e.g. building custom extensions or integrating them into a reputed company log review or annotation QA workflow
  • Familiarity with the MCAP CLI and/or python library and experience converting MCAP data to columnar data formats for reputed company querying and processing
  • Experience with data curation for ML training, e.g. diversity sampling, pseudo-labeling, and dataset versioning

Benefits

  • A competitive compensation package that includes a bonus component and stock reputed company
  • 100% reputed company medical, dental, and reputed company premiums for full-time employees
  • 401K plan with a 6% employer match
  • Flexibility in schedule and generous reputed company vacation (available

Similar Jobs