Staff Research Engineer (LLM Pre-Training) (m/f/d)
Core
Developing Large Language Models from scratch and deploying them into production environments for coding assistance across JetBrains products.
Role type
Staff Research Engineer (LLM Pre-Training)
Builds
In-house foundation models for writing and coding assistance integrated into JetBrains products
Domain
AI/ML, Large Language Models, Developer Tools
Deliverable
production ML models
Required skills
LLM training from scratch, distributed training of multi-billion parameter models, NLP and transformer-based approaches, deep learning frameworks (PyTorch), production ML system deployment, dataset collection and processing
Preferred skills
LLM inference frameworks (vLLM, DeepSpeed, TensorRT), LLM alignment techniques (RLHF/RLAIF), MLOps and CI/CD for ML, K8s and Kubeflow, scientific publications in NLP
Technologies
PyTorch, HuggingFace, Kubeflow, Weights & Biases, TeamCity, NVIDIA GPUs, Git, Python
Responsibilities
Train LLMs from scratch on a large GPU cluster, collect and process pre-training and fine-tuning datasets, work with stakeholders to convert business requirements into technical specifications, support and improve existing subsystems
Seniority
Staff, hands-on IC with strategic impact