CareerPlanSign in

Senior Applied ML Researcher - Video Apps

Cupertino, United States of America💼 Full-time🗓 2026-08-07 → 2026-09-28

Core

Design, train, and deploy state-of-the-art models for visual and audio understanding to enable intelligent systems that see, hear, and reason about the world.

Role type

Senior Applied ML Researcher (multimodal)

Builds

Deep neural networks for video, image, audio, and audio-visual tasks

Domain

Computer vision, audio signal processing, multimodal learning

Deliverable

production ML models

Required skills

deep neural networks, modern training workflows, computer vision, audio modeling, Python, PyTorch, linear algebra, probability, optimization

Preferred skills

PhD in CS/ML, publications in top-tier ML conferences, self-supervised learning, foundation model pre-training, open-source contributions, Objective-C, Swift

Technologies

PyTorch, Python

Responsibilities

Design and train deep neural networks for video, image, audio, and audio-visual tasks; Build models for audio-visual representation learning, cross-modal alignment, and fusion; Develop solutions for video understanding, temporal modeling, audio-visual event detection, and speech/sound/scene understanding

Seniority

Senior, hands-on IC

Sourced via apple · Listed on CareerPlan, which tracks 878,000+ jobs from 20+ sources.