Senior Applied ML Researcher - Video Apps
Core
Design, train, and deploy state-of-the-art models for visual and audio understanding to enable intelligent systems that see, hear, and reason about the world.
Role type
Senior Applied ML Researcher (multimodal)
Builds
Deep neural networks for video, image, audio, and audio-visual tasks
Domain
Computer vision, audio signal processing, multimodal learning
Deliverable
production ML models
Required skills
deep neural networks, modern training workflows, computer vision, audio modeling, Python, PyTorch, linear algebra, probability, optimization
Preferred skills
PhD in CS/ML, publications in top-tier ML conferences, self-supervised learning, foundation model pre-training, open-source contributions, Objective-C, Swift
Technologies
PyTorch, Python
Responsibilities
Design and train deep neural networks for video, image, audio, and audio-visual tasks; Build models for audio-visual representation learning, cross-modal alignment, and fusion; Develop solutions for video understanding, temporal modeling, audio-visual event detection, and speech/sound/scene understanding
Seniority
Senior, hands-on IC
