Senior Site Reliability Engineer (Remote)
Core
Operate and maintain Linux-based infrastructure, manage Kubernetes clusters, and design complex networking architectures for cloud computing solutions.
Role type
Senior Site Reliability Engineer (Infrastructure & Networking)
Builds
Production Kubernetes clusters, automated provisioning workflows, and observability stacks
Domain
Cloud Computing / Infrastructure Engineering
Deliverable
infrastructure
Required skills
Kubernetes production operations, Linux system administration, network engineering (VLANs, L2/L3 routing, VPNs), automation (Ansible, Bash, Python), observability stack management, virtualization technologies, distributed systems knowledge
Preferred skills
Service mesh implementation, Cloudflare API integration, GPU infrastructure management, security best practices, IT asset management
Technologies
Kubernetes, Debian, Ubuntu, Ansible, Bash, Python, GitOps, Prometheus, Grafana, Loki, ELK, Graylog, OpenStack, Proxmox, VMware, MAAS, VLANs, L2/L3 routing, VPNs
Responsibilities
Operate and maintain Linux-based infrastructure; Deploy, manage, and scale Kubernetes clusters; Design and maintain networking architecture; Implement automation for provisioning and operations; Lead incident response and escalation activities; Coordinate physical maintenance for hardware and data centers
Seniority
Senior, hands-on IC