Site Reliability Engineer - Comcast Technology Solutions
Core
Analyze and forecast system capacity, respond to incidents, optimize cloud infrastructure, and automate operations to ensure scalable, high-performance services for media and streaming customers.
Role type
Site Reliability Engineer (SRE)
Builds
Cloud Video Platform (CVP) and streaming services for media operators and content providers
Domain
Media & Entertainment / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Cloud platforms (AWS, GCP, Azure), Scripting (Python, Bash), Infrastructure-as-Code (Terraform, Ansible), Container orchestration (Kubernetes, Docker), Monitoring (Datadog, Splunk), Database performance tuning
Preferred skills
CI/CD pipeline management, Cloud cost optimization, Front-end development (React)
Responsibilities
Analyze and forecast system capacity requirements, Participate in incident response and post-incident reviews, Develop and maintain monitors and alerts, Optimize system performance through tuning, Develop and maintain disaster recovery plans, Identify opportunities for automation
Seniority
Mid-Senior, hands-on IC