SRE Systems Engineer (TS/SCI Clearance)
Core
Monitor customer-facing services, resolve critical incidents, and automate production issues to ensure Salesforce cloud availability.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Resilient cloud infrastructure and automated incident response systems for Salesforce's global customer base.
Domain
Cloud Infrastructure / SRE / AI CRM
Deliverable
production ML models | infrastructure
Required skills
Incident management (Sev0/Sev1), Unix CLI (Red Hat Enterprise Linux), TCP/IP networking, Python/Go scripting, AWS infrastructure provisioning, Root Cause Analysis (RCA), automation scripting
Preferred skills
Configuration management (Chef/Puppet), CI/CD pipelines (Jenkins/Bamboo/Spinnaker), Kubernetes management, Linux+/Red Hat/AWS certifications, Agile/DevOps practices
Technologies
AWS, Red Hat Enterprise Linux, Python, Go, Kubernetes, Jenkins, Bamboo, Spinnaker, Chef, Puppet
Responsibilities
Monitor and respond to Severity 0 and Severity 1 incidents; Automate detection and resolution of recurring production issues; Contribute to compliance, resiliency, and self-healing initiatives; Partner with and mentor team members on industry technology
Seniority
Senior, hands-on IC