Senior Site Reliability Engineer
Core
Own infrastructure for authentication and authorization services across production EKS clusters globally, integrating AI agents to accelerate engineering workflows.
Role type
Senior Site Reliability Engineer (Infrastructure & AI Integration)
Builds
Scalable, automated platform foundation for identity services on EKS
Domain
Cybersecurity, Identity and Access Management (IAM), Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Kubernetes, GitOps (Flux/ArgoCD), Infrastructure as Code (Terraform, Helm, Crossplane), Container orchestration (EKS), Monitoring (New Relic, Splunk, PagerDuty), CI/CD (CircleCI), Secrets management (SOPS, AWS KMS), AI agents/LLM tooling
Preferred skills
Python or Go scripting, Internal AI agent development, Identity protocols (OAuth 2.0, SAML, mTLS), Ping Identity products knowledge
Technologies
EKS, Flux, ArgoCD, Terraform, Helm, Crossplane, New Relic, Splunk, PagerDuty, CircleCI, SOPS, AWS KMS
Responsibilities
Drive deployment automation across production and test EKS clusters using GitOps; Take ownership of infrastructure requirements and coordinate the maintenance backlog; Integrate and extend AI agents to accelerate internal engineering workflows; Build and maintain a scalable, automated platform foundation across production and development environments; Contribute to identity-focused platform infrastructure supporting authentication and authorization services
Seniority
Senior, hands-on IC