
$139,000 – $166,800
Listed on Torc Robotics’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role involves building and maintaining the data infrastructure that powers Torc's autonomous truck software, including data pipelines, storage systems, and tooling for processing large-scale sensor logs from vehicles in the field. You'll work on a small team with high ownership, responsible for turning raw vehicle data into the curated datasets that machine learning and autonomy engineers depend on.
Our summary, not Torc Robotics’s wording. The full posting is on their site.
Skills this role names
- Amazon Athena
- Amazon Redshift
- Amazon S3
- Amazon Web Services (AWS)
- AWS CloudFormation
- AWS Glue
- Python
- SQL
- Terraform
Log in to see which of these are already on your profile.
What they ask for
Required
- Bachelor's degree in Computer Science, Computer Engineering, Software Engineering, Electrical Engineering, or related field with 4+ years data engineering experience, or Master's degree with 2+ years experience
- Strong proficiency in Python and SQL for production data pipelines
- Experience with cloud data infrastructure (AWS: S3, Glue, Athena, Redshift or equivalent)
- Experience with infrastructure-as-code tools (Terraform or CloudFormation)
- Understanding of data partitioning strategies and columnar storage formats
- Experience building and operating pipelines for time-series and binary data
- Ability to evaluate and integrate open-source tools appropriately
- Strong practices in data quality, monitoring, validation, and lineage tracking
- U.S. citizenship required
Nice to have
- Experience with autonomous vehicles, robotics, or sensor-driven autonomous systems
- Deep experience with Foxglove or Rerun including custom extensions or integration into log review workflows
- Familiarity with MCAP CLI and Python library and converting MCAP to columnar formats
- Experience with data curation for ML training including diversity sampling, pseudo-labeling, and dataset versioning