Software Engineer, II – Autonomy Data

Job not on LinkedIn

🔥 1 minute ago

⚔️ Virginia – Remote

info

💵 $139k - $166.8k / year

⏰ Full Time

🟢 Junior

🟡 Mid-level

🚰 Data Engineer

🦅 H1B Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Torc Robotics

Torc Robotics

501 - 1000 employees

Founded 2007

🚘 Automotive

📦 Logistics

🚗 Transport

Automotive • Logistics • Transport

Torc Robotics is an innovative company focused on commercializing self-driving trucks for long-haul freight transportation. As an independent subsidiary of Daimler Truck, the company is developing autonomous technology, primarily focusing on the Freightliner Cascadia. Torc is committed to safe transportation, continuously improving its solutions through rigorous testing and integration of industry-leading sensors. It collaborates with fleet management companies to deploy real-world autonomous solutions, aiming to lead the industry in autonomous trucking.

📋 Description

• Contribute to the design and organization of the autonomy program’s data lake, including schema definitions, partitioning strategy, and metadata indexing • Build and maintain reliable end-to-end pipelines ingesting high-bandwidth vehicle sensor logs into cloud storage • Implement data validation and integrity checks for corrupted information, missing sensors, and inconsistent calibration • Implement data retention, tiering, and lifecycle policies • Build tooling to query raw logs and produce curated training and evaluation datasets • Automate cost-effective pseudo-labeling workflows at ingest scale • Implement data quality and model performance metrics to direct labeling effort • Deploy and maintain visualization tooling for log review, annotation QA, and autonomy debugging • Integrate visualization tooling with the data lake to connect dataset entries or model failures to source logs • Define visualization panels and metrics with autonomy engineers • Build dashboards showing data coverage by terrain, operating environment, and geographic region • Establish and document data contracts between data services and model-training consumers • Partner with perception, planning, and embedded engineers across the data lifecycle • Help evolve data engineering standards, best practices, and tooling choices • Contribute to the data roadmap and surface findings to senior technical leadership

🎯 Requirements

• Bachelor’s degree in Computer Science, Computer Engineering, Software Engineering, Electrical Engineering, or a related field with 4+ years of data engineering experience, or a Master’s with 2+ years • Strong proficiency in Python and SQL • Ability to build production-quality data pipelines • Experience with cloud data infrastructure, preferably AWS S3, Glue, Athena, Redshift, or equivalent • Experience with infrastructure-as-code tools such as Terraform or CloudFormation • Understanding of data partitioning strategies and columnar storage formats such as Parquet and ORC • Experience building and operating pipelines processing time-series and binary data • Ability to evaluate and integrate open-source tooling • Experience implementing monitoring, validation, data quality, and lineage tracking • Only U.S. citizens are eligible for this role • Bonus: experience with autonomous vehicles, robotics, or sensor-driven autonomous systems • Bonus: deep experience with Foxglove or Rerun • Bonus: familiarity with MCAP CLI or Python library and MCAP-to-columnar conversion • Bonus: experience with ML data curation, diversity sampling, pseudo-labeling, and dataset versioning

🏖️ Benefits

• A competitive compensation package that includes a bonus component and stock options • 100% paid medical, dental, and vision premiums for full-time employees • 401K plan with a 6% employer match • Flexibility in schedule and generous paid vacation (available immediately after start date) • Company-wide holiday office closures • AD+D and Life Insurance • Potential sign-on payments, relocation, and other forms of compensation as part of the total compensation package

Apply Now

Similar Jobs

🔥 1 hour ago

NBCUniversal

10,000+ employees

📱 Media

Data Engineer building scalable AWS, Spark, and Databricks pipelines for Comcast Advertising’s multiscreen TV advertising technology. Designing data lakes, semantic layers, and self-service analytics platforms for large-scale users.

🔥 3 hours ago

LMI

1001 - 5000

📦 Logistics

🏥 Healthcare

🎖️ Defense

Data Engineer building secure, machine-learning-ready pipelines for LMI, a federal digital-solutions provider. Supporting SOCOM analytics, AI/ML products, integrations, and mission sustainment.

🔥 4 hours ago

Cummins Inc.

10,000+ employees

🏗️ Construction

💼 Consulting

🏥 Healthcare

Data Architect designing scalable cloud data platforms and ETL pipelines for Cummins, a global engine manufacturer. Optimizing Snowflake, Fabric, SAP, graph, and streaming integrations.

🇺🇸 United States – Remote

💵 $123k - $150.4k / year

💰 $75M Grant on 2024-07

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

🔥 5 hours ago

CVS Health

10,000+ employees

🏥 Healthcare

⚕️ Healthcare Insurance

🛒 Retail

Data Engineer building cloud and on-premise data solutions for CVS Health’s healthcare services. Processing payer files and migrating workloads across data platforms.

🔥 5 hours ago

Guidehouse

10,000+ employees

🏥 Healthcare

🎖️ Defense

📦 Logistics

Data Engineer building Palantir data pipelines and governance tools for Guidehouse’s government and defense clients. Applying industrial engineering and operations research to improve depot performance and modernization decisions.