Site Reliability Engineer
Core
Ensure availability and performance of customer-facing platform by managing infrastructure, deploying applications, and automating workflows.
Role type
Site Reliability Engineer (SRE)
Builds
Highly available Windows/Linux systems, CI/CD pipelines, and automated infrastructure
Domain
Industrial technology / Cloud infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
AWS, Terraform, Ansible, Kubernetes, Bash/PowerShell/Python, CI/CD tools, observability tools, networking, SQL Server
Preferred skills
Azure, Docker, WISA stack (Windows/IIS/SQL Server/ASP.net)
Technologies
AWS, Azure, Terraform, ARM Templates, Cloud Formation, GitHub Actions, Octopus, Ansible, Jenkins, Azure DevOps, New Relic, Application Insights, AppDynamics, DataDog, Kubernetes, Docker, SQL Server, IIS
Responsibilities
Manage and maintain highly available systems (Windows and Linux), analyze metrics and trends for scalability, automate workflows, create infrastructure as code, maintain data backups and disaster recovery plans, design and deploy CI/CD pipelines, adhere to security best practices, follow ITIL best practices
Seniority
Mid-Senior, hands-on IC