Research Engineer, FlexOlmo
Core
Designing and implementing scalable infrastructure, architectures, and training pipelines for next-generation large language models (LLMs) with a focus on Mixture-of-Experts (MoE) and long-context models.
Role type
Senior Research Engineer (LLM Infrastructure & Architecture)
Builds
Scalable machine learning pipelines, training systems, and open-source tools for foundation models.
Domain
Artificial Intelligence / Large Language Models / Deep Learning
Deliverable
production ML models
Required skills
Building ML infrastructure, optimizing training and inference, Python, PyTorch/Jax/Tensorflow, cloud compute (AWS), containerization (Docker), debugging performant systems, data preprocessing/transformation, model evaluation and monitoring.
Preferred skills
Advanced degree in CS/ML/NLP, open-source contributions, operating models at scale in production, HPC experience.
Technologies
Python, PyTorch, Jax, Tensorflow, AWS, Docker
Responsibilities
Building infrastructure to facilitate next-generation LLM research, optimizing training and inference for language models, triaging between experiments, bridging research and product, bringing software engineering best practices to research, supporting open-source community, releasing contributions as open source software and technical reports.
Seniority
Senior, hands-on IC