Site Reliability Engineer - Neovest
Core
Design, develop, test, and deliver availability, reliability, scalability, and related application solutions for mission-critical systems in Corporate and Investment Banking.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Automated CI/CD pipelines, infrastructure as code, and network as code for applications and platforms
Domain
Financial Services / Cloud Infrastructure / Site Reliability Engineering
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Python, Java/Spring Boot, .Net, Azure, observability (white-box/black-box monitoring, SLO alerting, telemetry), CI/CD tools (Jenkins, GitLab, Terraform), containers (ECS, Kubernetes, Docker), networking troubleshooting
Preferred skills
Experience with troubleshooting common networking technologies, ability to present information clearly, proactive issue identification
Technologies
Azure, CI/CD, Datadog, Docker, Dynatrace, GitLab, Grafana, Jenkins, Kubernetes, Prometheus, Python, Splunk, Spring Boot, Terraform, ASP.NET
Responsibilities
Collaborate with software engineers to design and implement deployment approaches using automated CI/CD pipelines; Implement infrastructure as code, configuration as code, and network as code; Use service level indicators and objectives to address issues proactively; Help drive adoption of site reliability engineering best practices across the team
Seniority
Senior, hands-on IC