Site Reliability Engineer with DevOps expertise
Core
Site Reliability Engineer responsible for cloud infrastructure reliability, incident resolution, and automation of CI/CD pipelines.
Role type
Senior Site Reliability Engineer (DevOps)
Builds
Cloud-based solutions, automated deployment pipelines, and high-availability infrastructure
Domain
Cloud Infrastructure (AWS), DevOps, Data Processing
Deliverable
production ML models | infrastructure
Required skills
Linux system administration, AWS services (Lambda, S3, Glue, Route 53, IAM), Kubernetes, SQL databases, ETL tools (Control-M, Informatica), CI/CD pipelines, scripting (Python, Java, Bash, PowerShell), monitoring (Datadog, Splunk), SLI/SLO metrics
Preferred skills
IT Service Management (Incident, Change, Problem Management), creative problem solving in unstructured environments
Technologies
AWS Lambda, Bash, CI/CD, Cloud, Datadog, ETL, Git, Groovy, IAM, Informatica, JIRA, Java, Jenkins, Kubernetes, Linux, Maven, MySQL, Oracle, Python, SQL, Sonar, Splunk
Responsibilities
Collaborate with development teams to execute agile workflows, resolve high-impact incidents, develop and transition to cloud-based solutions, and utilize creative thinking to identify secure solutions
Seniority
Senior, hands-on IC