Software Engineer, II - Data Engineering
Core
Building secure, scalable data management solutions (ingestion, ETL, storage) to power analytics, simulation, and ML training for autonomous truck fleets.
Role type
Mid-level IC data engineering software engineer
Builds
AWS-native data pipelines, data lake infrastructure, and cloud-based solutions for ML training
Domain
Autonomous driving / Robotics / Cloud Data Engineering
Deliverable
production ML models | infrastructure
Required skills
Linux system administration, Bash scripting, Docker containerization, AWS serverless services, Infrastructure as Code (Terraform), CI/CD pipelines, Data warehousing, NoSQL databases, DAG workflow orchestration
Preferred skills
Databricks, Ray for ML scaling, Python data analysis libraries (pandas, numpy), Robotics data acquisition patterns
Technologies
AWS (Lambda, Step Functions, Glue, Athena, Batch, EventBridge), Linux, Docker, Terraform, Python, Jenkins, GitHub Actions, Datadog, React
Responsibilities
Create robust pipelines for massive daily data volumes from vehicle fleets; Build and support scalable pipelines for the Data Factory; Scale the data lake via distributed storage; Promote data integrity through validation and governance; Collaborate with perception and planning teams; Participate in on-call rotation for deployed systems
Seniority
Mid-level, hands-on IC