Site Reliability Engineer (w/m/d)
Core
Ensuring stable operation of business-critical software solutions along the sales value chain, reducing outages, and maintaining modern, secure infrastructure.
Role type
Site Reliability Engineer (SRE)
Builds
Modern infrastructure, automation tools, and monitoring solutions for sales process applications.
Domain
Retail (optical industry) + Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Infrastructure as Code, Docker, Kubernetes, Linux, TCP/IP, CI/CD pipelines, scripting (Bash/PowerShell), monitoring implementation, incident management, documentation, on-call support
Preferred skills
Terraform, Ansible, Puppet, AWS, Azure, OCI, process automation tools (GitHub Actions, Concourse, Jenkins, TeamCity)
Responsibilities
Ensure stable operation of business-critical software solutions; identify and resolve errors, outages, and performance bottlenecks; develop tools and automations to improve infrastructure reliability; implement and optimize monitoring solutions; build and operate modern infrastructure with IaC and CI/CD; collaborate with development teams on best practices; create emergency plans and document system configurations; provide on-call support.
Seniority
Mid-level, hands-on IC