DevOps & SRE Engineer
Core
Design, build, and maintain scalable, high-availability distributed systems and infrastructure platforms for production services.
Role type
Senior DevOps & SRE Engineer
Builds
Container clusters, CI/CD pipelines, monitoring/alerting systems, and self-service operational tools.
Domain
Cloud-native infrastructure and distributed systems
Deliverable
production ML models | infrastructure
Required skills
Kubernetes, Docker, CI/CD pipelines, observability tools, Python/Shell scripting, public cloud platforms (AWS/Azure/GCP), distributed systems architecture
Preferred skills
Service Mesh architectures, Cilium CNI, eBPF technologies, network security, load balancing, traffic management
Technologies
Kubernetes, Docker, Kafka, Redis, Elasticsearch, Nginx, MySQL, AWS, Azure, GCP, Shell, Python
Responsibilities
Manage and maintain container clusters and open-source component clusters; Design and enhance infrastructure operation platforms; Ensure maximum uptime through proactive monitoring and incident response; Lead development of automated operations and maintenance systems; Establish best practices for infrastructure code and configuration management; Continuously optimize service architecture and deployment strategies.
Seniority
Mid-level, hands-on IC