Machine Learning Operations Engineer
Core
Design and implement a self-service MLOps platform for data ingestion, feature engineering, model training, validation, deployment, and monitoring in a cloud environment.
Role type
Senior MLOps Engineer (Infrastructure & Platform)
Builds
MLOps platform, CI/CD pipelines, IaC solutions, monitoring dashboards
Domain
Cloud Infrastructure, Machine Learning Operations, DevOps
Deliverable
production ML models | infrastructure
Required skills
MLOps lifecycle management, Infrastructure as Code (Terraform), CI/CD pipeline design, Python scripting, Bash scripting, Kubernetes, Azure services, Microsoft Entra ID, GNU/Linux administration
Preferred skills
Ansible, Agile/SAFe methodologies, stakeholder communication
Technologies
Terraform, GitHub Actions, Kubernetes, Azure, Microsoft Entra ID, Python, Bash, GNU/Linux, Ansible
Responsibilities
Collaborate on MLOps strategy and technical design, develop IaC solutions for cloud resource provisioning, implement security measures for ML artifacts, create monitoring dashboards and alerts, maintain documentation for MLOps processes, participate in code reviews and best practice development
Seniority
Senior, hands-on IC