Senior Engineer
Core
Own 24x7 health and operational readiness for Skylo's hybrid production infrastructure, serving as the L3 escalation authority for cloud infrastructure incidents.
Role type
Senior Site Reliability Engineer (Infrastructure)
Builds
Hybrid production infrastructure including GCP, Kubernetes, bare-metal systems, storage, databases, and observability pipelines.
Domain
Cloud Infrastructure / Satellite Communications
Deliverable
infrastructure
Required skills
Kubernetes, GCP, PostgreSQL, Redis, GitOps (ArgoCD/Flux), Prometheus, Grafana, Terraform, Linux, Network Troubleshooting, SLO/SLI definition, Root Cause Analysis
Preferred skills
Telecom/NTN infrastructure, Ceph, KubeVirt, BGP, VXLAN, Go, Python, FinOps, Infrastructure Certifications
Technologies
GCP, Kubernetes, Prometheus, VictoriaMetrics, Grafana, OpenTelemetry, Loki, ELK, PostgreSQL, Redis, ArgoCD, Flux CD, Helm, Terraform, Ansible
Responsibilities
Operate and improve observability pipelines; Maintain database replication and cluster reliability; Define SLOs and capacity plans; Lead root-cause analysis and post-incident actions; Mentor Senior NREs and collaborate with platform and security teams.
Seniority
Senior, hands-on IC