Senior Site Reliability Engineer
Core
Design, build, and evolve scalable, resilient AWS cloud infrastructure and CI/CD pipelines for high-volume diagnostic workflows serving veterinarians and reference laboratories.
Role type
Senior Site Reliability Engineer (IC)
Builds
AWS Serverless platforms, automated deployment pipelines, and observability systems for IDEXX Reference Lab operations.
Domain
Veterinary diagnostics / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
AWS Serverless architecture, Terraform, CloudFormation, GitHub Actions, CI/CD pipeline design, distributed system troubleshooting, security vulnerability remediation, incident response, root cause analysis, release engineering governance, observability (metrics, logging, tracing), Kubernetes (implied by standard SRE context but not explicitly listed as required, sticking to text: AWS Lambda, SQS, SNS, EventBridge, DynamoDB, S3, AuroraDB, Cloudfront), Maven, Git workflows, OAuth2, OpenID Connect, Azure Entra ID.
Preferred skills
Kotlin or Java development, NoSQL and relational databases (PostgreSQL, DynamoDB), SLA/SLO/SLI management, distributed tracing (AWS X-Ray, OpenTelemetry), AI tools for automation.
Technologies
AWS (Lambda, SQS, SNS, EventBridge, DynamoDB, S3, AuroraDB, Cloudfront), Terraform, CloudFormation, GitHub Actions, Maven, Git, Azure Entra ID, PostgreSQL, DynamoDB, JFrog Artifactory, AWS X-Ray, OpenTelemetry.
Responsibilities
Own CI/CD pipeline architecture and governance; modernize deployment pipelines for AWS Lambda services; design and implement disaster recovery and high availability; build end-to-end observability and alerting; lead incident response and postmortems; govern release processes and deployment approvals; remediate security vulnerabilities and embed secure practices.
Seniority
Senior, hands-on IC