Senior Site Reliability Engineer
Core
Lead initiatives to transform IT Compute Core architecture and build new infrastructure service offerings across on-premises and cloud environments.
Role type
Principal Staff Site Reliability Engineer (Infrastructure)
Builds
Core infrastructure services including DNS, NTP/PTP, DHCP, and LDAP at global scale
Domain
Cloud and on-premises compute infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Large-scale compute platform engineering, automation, containerization architectures, distributed-systems infrastructure, Go/Python, Linux kernel internals, network protocols (VLAN, VXLAN, SDN, BGP, Anycast), infrastructure as code, configuration management
Preferred skills
eBPF, XDP, SR-IOV, DPU capabilities, microservices architecture, enterprise-wide capacity planning
Technologies
Terraform, DNS, LDAP, eBPF, XDP, SR-IOV, DPU
Responsibilities
Design, scale, and deploy core infrastructure services; define and implement service-efficiency metrics; develop tools for data analysis, performance profiling, and monitoring; collaborate with leadership to develop IT products and services
Seniority
Principal, hands-on IC with technical leadership