Site Reliability Engineer | Core Software Engineering Services Team
Core
Build and maintain the internal tool stack and monitoring infrastructure to support daily business operations and service health for colleagues across the company.
Role type
Site Reliability Engineer (SRE)
Builds
Internal CI/CD pipelines, automation servers, and monitoring solutions for engineering teams
Domain
Software Engineering Services / DevOps
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
CI/CD pipeline management, container orchestration, Linux/Windows system administration, scripting automation, monitoring and alerting, root cause analysis, capacity planning
Preferred skills
Infrastructure as Code, cloud platform configuration, distributed team collaboration
Technologies
Jenkins, GitHub Actions, Docker, Kubernetes, Linux Shell, Python, Zabbix, Grafana, Kibana, Gerrit, Git, Jira, Confluence, Ansible, Zuul, Azure DevOps, Openstack
Responsibilities
Support CI/CD pipelines to automate software delivery; Automate routine operational tasks and system maintenance; Administer and optimize automation servers; Implement and enhance monitoring and testing solutions; Perform technical root cause analysis and corrective actions; Participate in capacity planning and scaling strategies; Provide mentorship to junior engineers; Collaborate across geographically distributed teams
Seniority
Mid-Senior, hands-on IC