Software Engineer, Site Reliability
Core
Build and harden mission-critical infrastructure for a medical AI platform used by healthcare providers worldwide, applying SRE principles to improve systemic health, performance, and efficiency.
Role type
Site Reliability Engineer (SRE)
Builds
Mission-critical infrastructure, data platforms, and observability systems for a global medical AI platform
Domain
Healthcare / Medical AI / Cloud Infrastructure
Deliverable
infrastructure
Required skills
SRE mindset, incident response, observability, service objective definition, infrastructure hardening, toil reduction
Preferred skills
end-to-end ownership, autonomous execution, 0->1 and 1->1000 scaling experience
Technologies
null
Responsibilities
Ensure seamless and effective on-call processes and incident response; define robust service objectives across services and data platforms; improve systemic health, performance, and efficiency of the platform
Seniority
Mid-Senior, hands-on IC