Software Engineer - Kafka
Core
Building large-scale data platforms and machine learning infrastructure to process petabytes of multimodal sensor data from real-world driving scenarios for AI research and autonomy stack development.
Role type
Senior backend software engineer (data engineering & ML infrastructure)
Builds
Central data and machine learning platform supporting data ingestion, processing, storage, labeling infrastructure, and workflow orchestration for the autonomy stack.
Domain
Autonomous driving, physical AI, data engineering
Deliverable
production ML models | infrastructure
Required skills
Backend engineering, distributed computing frameworks, data processing systems, data curation and tagging, cross-functional collaboration
Preferred skills
Apache Spark, Apache Hudi, Trino, Apache Kafka, Flyte, data lake architectures, streaming systems, workflow orchestration
Technologies
Apache Spark, Apache Hudi, Trino, Apache Kafka, Flyte, Kubernetes, Python, Golang, Java
Responsibilities
Design and build large-scale data platforms handling petabytes of sensor data; work on data curation and tagging platforms for dataset discovery and labeling; build high-performance data processing systems to transform raw sensor data into training-ready formats.
Seniority
Senior, hands-on IC