CareerPlanSign in

Senior Engineer

USA💼 Full-time🗓 2026-09-24 → 2026-09-26

Core

Own 24x7 health and operational readiness for Skylo's hybrid production infrastructure, serving as the L3 escalation authority for cloud infrastructure incidents.

Role type

Senior Site Reliability Engineer (Infrastructure)

Builds

Hybrid production infrastructure including GCP, Kubernetes, bare-metal systems, storage, databases, and observability pipelines.

Domain

Cloud Infrastructure / Satellite Communications

Deliverable

infrastructure

Required skills

Kubernetes, GCP, PostgreSQL, Redis, GitOps (ArgoCD/Flux), Prometheus, Grafana, Terraform, Linux, Network Troubleshooting, SLO/SLI definition, Root Cause Analysis

Preferred skills

Telecom/NTN infrastructure, Ceph, KubeVirt, BGP, VXLAN, Go, Python, FinOps, Infrastructure Certifications

Technologies

GCP, Kubernetes, Prometheus, VictoriaMetrics, Grafana, OpenTelemetry, Loki, ELK, PostgreSQL, Redis, ArgoCD, Flux CD, Helm, Terraform, Ansible

Responsibilities

Operate and improve observability pipelines; Maintain database replication and cluster reliability; Define SLOs and capacity plans; Lead root-cause analysis and post-incident actions; Mentor Senior NREs and collaborate with platform and security teams.

Seniority

Senior, hands-on IC

Sourced via codingjobboard · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.