Mid/Senior Site Reliability Engineer (US Federal)
Core
Design, implement, test, deploy, and maintain automation infrastructure for configuration management and service deployment to support the SDLC of application developers.
Role type
Mid/Senior Site Reliability Engineer (Infrastructure & Automation)
Builds
Automation infrastructure, binary artifact management, software distribution, source control systems, and build tools.
Domain
Cloud Infrastructure, DevOps, Site Reliability Engineering
Deliverable
production ML models | product features | infrastructure
Required skills
Configuration management (Chef, Ansible, Terraform), Container orchestration (Docker, Kubernetes), Cloud platforms (AWS, GCP, Azure), Linux administration, Scripting (Python, Ruby, Golang), CI/CD pipelines, Incident response and triage, Reliability Engineering principles, Public Cloud Networking, Source Code Management (SCM) tools.
Preferred skills
JFrog Artifactory administration, DoD 8570/8140 compliance (IAT Level II), Security certifications (CompTIA CySA+, GICSP, CASP+), TS/SCI clearance.
Responsibilities
Define, design, implement, test, deploy, and maintain automation infrastructure; Improve operational efficiency by automating processes; Build platforms for self-service production interaction; Respond to production monitoring, triage, fix, and resolve incidents; Participate in on-call rotation.
Seniority
Mid-level (3+ years) / Senior-level (5+ years), hands-on IC