Senior Machine Learning Engineer
Core
Design, build, and maintain MLOps systems (microservices, APIs, orchestration) to scale workflows for large data, distributed systems, and LLM/generative AI development.
Role type
Senior Machine Learning Engineer (MLOps/Infrastructure)
Builds
Scalable MLOps infrastructure, microservices, queuing systems, and orchestration workflows for LLM and generative AI models.
Domain
Cloud computing, distributed systems, and AI infrastructure.
Deliverable
infrastructure
Required skills
Python, Linux, Kubernetes, Terraform, CI/CD pipelines, Prometheus, Grafana, OpenTelemetry, distributed system design, cloud computing (AWS/GCP)
Preferred skills
Systems thinking, test-driven development, mentoring junior engineers
Technologies
Python, Kubernetes, Kafka, Prometheus, Grafana, Terraform, AWS, GCP, OpenTelemetry
Responsibilities
Design and maintain MLOps systems including microservices and orchestration workflows; implement observability tools for production ML systems; review architecture plans for scalability; mentor junior and mid-level engineers.
Seniority
Senior, hands-on IC with mentorship responsibilities