Research Engineer, FlexOlmo Seattle, WA View job
Core
Design and implement scalable ML infrastructure and training pipelines for next-generation large language model architectures, focusing on Mixture-of-Experts, long-context models, and flexible data usage.
Role type
Senior IC research engineer (LLM infrastructure & training)
Builds
Scalable machine learning pipelines, training systems, and open-source model releases
Domain
Artificial Intelligence / Large Language Models / Deep Learning
Deliverable
production ML models
Required skills
Python, PyTorch/Jax/Tensorflow, deep learning, natural language processing, software engineering, cloud compute (AWS), containerization (Docker)
Preferred skills
HPC operations, open-source contributions, production-scale model deployment
Technologies
PyTorch, Jax, Tensorflow, AWS, Docker
Responsibilities
Build infrastructure for next-gen LLM research; optimize training and inference; triage experiments; bridge research to product; release open-source software and technical reports
Seniority
Senior, hands-on IC