Senior Site Reliability Engineer
Core
Ensure scalability, reliability, and performance of cloud infrastructure and services for a global agricultural technology platform.
Role type
Senior Site Reliability Engineer (SRE)
Builds
SaaS applications for growers, agronomists, and ag retailers
Domain
AgTech / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Linux, Bash, Cloud platforms (AWS/GCP/Azure), Python/Go/Ruby, Terraform, Docker, Kubernetes, CI/CD pipelines, Git, Observability tools (Datadog/New Relic/Splunk)
Preferred skills
Service mesh (Envoy/Istio), NATS, AI/LLM tooling
Technologies
AWS, GCP, Azure, Terraform, Docker, Kubernetes, EKS, Buildkite, Git, Datadog, New Relic, Prometheus, Splunk, Envoy, Istio, NATS
Responsibilities
Lead infrastructure projects and perform high-risk maintenance; Resolve incidents and participate in on-call rotations; Mentor team members in SRE practices; Identify system scalability issues and drive architectural improvements; Maintain Service Level Indicators (SLIs) and SLOs; Promote automation to reduce operational overhead.
Seniority
Senior, hands-on IC with leadership responsibilities