Senior Site Reliability Engineer, NetBox Delivery
Core
Senior SRE responsible for the reliability, observability, and delivery automation of a network infrastructure platform (NetBox) at scale.
Role type
Senior Site Reliability Engineer (Delivery & Platform)
Builds
Network infrastructure platform (NetBox) and its delivery pipeline for enterprise customers
Domain
Network Infrastructure / Cloud Engineering
Required skills
Python, Django, PostgreSQL, Kubernetes, Terraform, AWS, CI/CD, Observability, Incident Response, Supply-chain security
Preferred skills
NetBox ecosystem, Open-source contributions, B2B startup experience, SLSA/cosign expertise
Technologies
AWS (EC2, VPC, IAM, RDS), Kubernetes, Helm, GitHub Actions, ArgoCD, FluxCD, Terraform, Prometheus, Grafana, Django, PostgreSQL, Python, cosign, Sigstore, SLSA
Responsibilities
Own and improve software build and release pipelines; Design reliable release handoffs between core engineering and deployment teams; Improve application performance and production reliability; Build and maintain comprehensive observability; Serve as escalation point for performance and reliability issues; Strengthen software supply-chain security; Support compliance initiatives; Participate in on-call rotation and lead incident response; Conduct postmortems and drive engineering improvements; Drive technical initiatives across team boundaries; Help shape engineering team practices and operating model.