Site Reliability Engineer Intern
Core
Ensure stability, performance, and reliability of Copart's critical applications and global data center infrastructure through monitoring, troubleshooting, and automation.
Role type
Site Reliability Engineer Intern
Builds
Monitoring and automation tools for global data centers and application infrastructure
Domain
Online vehicle auctions / Cloud infrastructure
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Linux, Windows, scripting, automation, observability tools, incident management, root cause analysis, SOP creation
Preferred skills
Python, Ansible, Datadog, Kubernetes, VMware vSphere, AWS, GCP, AI tools
Technologies
Python, Ansible, Datadog, Kubernetes, VMware vSphere, AWS, GCP
Responsibilities
Proactively monitor global data centers and application infrastructure; Design and optimize monitoring/automation tools; Support incident management including triage and root cause analysis; Develop analytical reporting capabilities; Create and maintain SOPs and system diagrams; Coordinate periodic failover testing
Seniority
Intern