Skip to main content
Torc

Software Engineer II - Autonomy Data

RemoteUnited States only
Published
Role
Data Engineering
Experience
Mid
Employment
Full-time
$139k–$166.8k/yr
Check eligibility

Open to US only. Set where you work from to check your eligibility.

No BS summary

Level 2 autonomy data engineer with 4+ years of data engineering experience, or 2+ with a master's. Must be a U.S. citizen and strong in Python, SQL, cloud data infrastructure, IaC, and production data pipelines for time-series/binary data.

Core skills

PythonSQLAWS

Required skills

S3Glue AthenaRedshiftTerraformCloudFormationParquetORC

Optional skills

FoxgloveRerunMCAP CLIMCAP Python library

What you'll do

  • Contribute to the design and organization of the program’s data lake, including schema definitions, partitioning strategy, and metadata indexing.
  • Build and maintain end-to-end pipelines that ingest high-bandwidth sensor logs from vehicles into cloud storage with high reliability and tolerance of ad-hoc and intermittent connectivity mechanisms.
  • Implement data validation and integrity checks to detect corrupted information, missing sensors, and inconsistent calibration before downstream processing.
  • Implement retention, tiering, and lifecycle policies for data to balance storage costs with development value.
  • Build tooling to query raw logs to produce curated training and evaluation datasets.
  • Build automation to run cost-effective pseudo-labeling workflows at the scale of data ingest.
  • Implement data quality and model performance metrics used to direct labeling effort toward the highest-value examples.
  • Deploy and maintain data visualization tooling to support log review, annotation QA, and autonomy debugging workflows.
  • Build integrations between visualization tooling and the data lake so engineers can navigate from a dataset entry or model failure directly to the origin log data.
  • Work with autonomy engineers to define and surface custom visualization panels and implement metrics for analyzing unstructured operating environments.
  • Build dashboards that give autonomy engineers visibility into data coverage by terrain type, operating environment, and geographic region.
  • Establish and document data contracts between data services and model training consumers.
  • Partner with perception, planning, and embedded engineers across the data lifecycle, from shaping logging schemas and collection triggers to defining dataset interfaces for model training and evaluation.
  • Follow and help evolve data engineering standards, best practices, and tooling choices for an innovative and fast-paced team.
  • Contribute to the data roadmap and surface findings to senior technical leadership.

What they require

  • Bachelor’s degree in Computer Science, Computer Engineering, Software Engineering, Electrical Engineering, or a related field with 4+ years of data engineering experience, or a Master’s with 2+ years.
  • Strong proficiency in Python and SQL, with demonstrated ability to build production-quality data pipelines.
  • Experience with cloud data infrastructure, AWS preferred: S3, Glue Athena, Redshift, or equivalent, and infrastructure-as-code tools such as Terraform or CloudFormation.
  • Solid understanding of data partitioning strategies and columnar storage formats such as Parquet or ORC.
  • Experience building and operating data pipelines that process time-series and binary data.
  • Proven ability to evaluate and integrate open-source tooling when appropriate versus building from scratch.
  • Good instincts for delivering data quality through first-class implementations of monitoring, validation, and lineage tracking.
  • Only U.S. citizens are eligible for this role because the position requires access to information and systems restricted under U.S. law.
  • Preferred: Experience with autonomous vehicles, robotics, or other sensor-driven autonomous systems.
  • Preferred: Deep experience with Foxglove or Rerun beyond basic playback, such as building custom extensions or integrating them into a structured log review or annotation QA workflow.
  • Preferred: Familiarity with the MCAP CLI and/or Python library and experience converting MCAP data to columnar data formats for further querying and processing.
  • Preferred: Experience with data curation for ML training, such as diversity sampling, pseudo-labeling, and dataset versioning.

Benefits

  • Competitive compensation package that includes a bonus component and stock options.
  • 100% paid medical, dental, and vision premiums for full-time employees.
  • 401K plan with a 6% employer match.
  • Flexibility in schedule and generous paid vacation available immediately after start date.
  • Company-wide holiday office closures.
  • AD+D and Life Insurance.
  • Corporate bonus and stock option plan.
  • Depending on the position offered, sign-on payments, relocation, and other forms of compensation may be provided.
  • Full range of medical, financial, and/or other benefits.

Torc develops autonomous driving software for automated trucks and is part of the Daimler family. It has been a leader in autonomous driving since 2007.

Autonomous VehiclesEnterpriseville-torcy.fr/

Details

Visa sponsorshipNo
Apply routeGreenhouse
$139k–$166.8k/yr