Senior Site Reliability Engineer
Core
Build, deploy, and operate critical compute software for end-to-end satellite imaging operations in customer on-premises and cloud environments.
Role type
Senior Site Reliability Engineer (Infrastructure)
Builds
Next-generation Constellation as a Service platform supporting on-premises and cloud deployments
Domain
Space technology / Satellite operations / Cloud infrastructure
Deliverable
production ML models | infrastructure
Required skills
Cloud-native infrastructure, Kubernetes (Talos, RKE2, Proxmox, k3s), Infrastructure as Code (Terraform, Ansible, Helm, Kustomize), CI/CD (Jenkins, GitLab CI/CD, Argo CD, CircleCI), Distributed systems troubleshooting, Python, Bash, Resource optimization, Cluster tuning
Preferred skills
CUDA-based GPU programs, Security expertise (zero-trust, air-gapped environments), Hardware and network level implications
Technologies
Kubernetes, Terraform, Ansible, Helm, Kustomize, Jenkins, GitLab CI/CD, Argo CD, CircleCI, Prometheus, Grafana, OpenTelemetry, Alloy
Responsibilities
Build and deploy computing services in customer environments, Architect novel systems for air-gapped deployments, Clarify requirements from cross-functional stakeholders, Manage operations (deployments, orchestration, documentation), Scale architecture while ensuring availability, Resolve edge cases and write tests, Participate in on-call rotations
Seniority
Senior, hands-on IC
