Senior ML Engineer
Core
Architect and maintain the productionization layer of Invoca's ML stack, focusing on model serving, inference optimization, and agentic AI workflows for marketing and contact center teams.
Role type
Senior ML Engineer (MLOps & Inference Infrastructure)
Builds
Production-grade APIs, scalable inference infrastructure, and fine-tuned SLM/LLM models for conversation intelligence.
Domain
AI/ML, Contact Center, Marketing Technology
Deliverable
production ML models | product features | infrastructure
Required skills
MLOps (CI/CD, model versioning), Python, PyTorch, HuggingFace Transformers, SLM/LLM fine-tuning (LoRA, QLoRA, PEFT), inference optimization (quantization, batching), Triton Inference Server, Baseten, Kubernetes, GPU infrastructure, API development, model monitoring/evaluation
Preferred skills
RLHF, preference training
Technologies
Triton, Baseten, vLLM, TGI, SageMaker, Vertex AI, MLflow, Braintrust, Kubernetes
Responsibilities
Lead end-to-end MLOps and productionization; Design and optimize SLM/LLM deployment; Fine-tune language models for NLP applications; Evolve ML infrastructure; Collaborate across teams to build foundational ML systems.
Seniority
Senior, hands-on IC