Helix AI Engineer, Video Pretraining
Core
Developing large-scale video foundation models trained on real-world and robot-collected data to enable perception, prediction, and embodied reasoning for autonomous humanoid robots.
Role type
Senior IC machine-learning engineer (video pretraining)
Builds
Large-scale video foundation models and data pipelines for embodied AI systems
Domain
Robotics + Computer Vision + Multimodal AI
Deliverable
production ML models
Required skills
Large-scale model training, deep learning architectures for video/vision, dataset curation, distributed training systems, Python, PyTorch, scalable system engineering
Preferred skills
Frontier video models, video diffusion/autoregressive modeling, robotics/embodied AI, publication record in ML/CV
Technologies
PyTorch, transformer-based models, diffusion-based models, large GPU clusters
Responsibilities
Design and train large-scale video foundation models on diverse datasets, develop pretraining strategies for temporal dynamics, build models for transferable representations, explore video understanding architectures, implement efficient data pipelines, optimize model performance, collaborate with agent/robot learning teams, design evaluation frameworks
Seniority
Senior, hands-on IC