Site Reliability Engineer (w/m/d)
Core
Ensuring stable operation of business-critical software solutions along the sales value chain, reducing outages, and maintaining modern, secure infrastructure.
Role type
Site Reliability Engineer (IC)
Builds
Reliable software infrastructure, automation tools, and monitoring solutions for the sales process.
Domain
Retail (optical industry) + Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Infrastructure as Code, Docker, Kubernetes, Linux, TCP/IP, CI/CD pipelines, scripting (Bash/PowerShell), monitoring implementation, incident management, system documentation
Preferred skills
Terraform, Ansible, Puppet, AWS, Azure, OCI, process automation tools (GitHub Actions, Concourse, Jenkins, TeamCity)
Responsibilities
Ensure stable operation of business-critical software solutions; identify and resolve errors, outages, and performance bottlenecks; develop tools and automations to improve infrastructure reliability; implement and optimize monitoring solutions; build and operate modern infrastructure with IaC and CI/CD; collaborate with development teams on best practices; create emergency plans and document system configurations.
Seniority
Mid-level, hands-on IC