Platform Engineer (Infrastructure & DevOps)
Core
Design, deploy, and operate infrastructure for AI and software products, ensuring reliable CI/CD pipelines, backend services, and developer workflows.
Role type
Mid-level Platform Engineer (Infrastructure & DevOps)
Builds
CI/CD pipelines, worker machines, backend services, monitoring systems, and hybrid cloud/on-premise infrastructure
Domain
Software development, AI/ML operations, Infrastructure as Code
Deliverable
production ML models | product features | infrastructure
Required skills
Linux administration, Docker, Kubernetes, CI/CD systems, scripting (Python/shell), networking, secrets management, observability
Preferred skills
GitOps, Helm, GPU server management, Python backend services, PostgreSQL, Redis, Grafana, Prometheus, Loki, Tempo, self-hosted GitHub runners, distributed storage (Ceph, MinIO), hybrid cloud deployments
Responsibilities
Design and maintain infrastructure for AI and software products; Own CI/CD pipelines and deployment workflows; Operate backend services, databases, and supporting infrastructure; Improve observability through logs, metrics, and incident investigation; Manage secrets, service accounts, and network access; Support hybrid infrastructure (cloud and on-premise); Manage computational resources for ML training and batch processing; Build backup and recovery procedures; Collaborate with ML and software engineers to improve system reliability and scalability; Document infrastructure decisions and operational procedures
Seniority
Mid-level, hands-on IC