Data Science/AI ML Engineer
Core
Design and implement specialized Large Language Models (LLMs) and efficient sparse architectures, translating research into production-grade AI systems for enterprise customer experience orchestration.
Role type
Senior IC AI/ML Engineer (LLM architecture & production)
Builds
Specialized LLMs, reusable Python libraries, and GPU-accelerated cloud infrastructure for model serving
Domain
Enterprise AI / Large Language Models / Cloud Infrastructure
Deliverable
production ML models
Required skills
Transformer architectures, attention mechanisms, model internals, Python, PyTorch, HuggingFace Transformers, AWS SageMaker, EC2 GPU instances, S3, CloudFormation, CI/CD pipelines, unit testing
Preferred skills
Sparse architectures, expert routing, quantization (GPTQ, AWQ), distributed inference, Terraform, conversational AI, agentic AI systems
Technologies
PyTorch, HuggingFace, AWS SageMaker, AWS EC2, AWS S3, AWS CloudFormation, Jenkins, Git
Responsibilities
Design and implement specialized LLMs with focus on model weights and efficient sparse architectures; Participate in the full lifecycle from research spike to deployable artefact; Build and maintain evaluation frameworks to compare model variants; Support and extend cloud infrastructure for experimentation and model serving; Engage with the open-source ML ecosystem; Provide guidance to junior team members
Seniority
Senior, hands-on IC