Manager TAG and Encompass SRE
Core
Lead a global team of Site Reliability Engineers to drive system reliability, incident response, and automation for WEX's complex payment and financial systems.
Role type
Mid-level SRE Manager (hands-on IC with people leadership)
Builds
Resilient, observable, and efficient systems for B2B payments, expense reimbursements, and virtual card creation
Domain
Fintech / Payments / Cloud Infrastructure
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Team leadership and mentorship, incident response, automation scripting (Python/Bash/Go), monitoring and logging (Grafana/ELK/Splunk), container orchestration (Docker/Kubernetes), CI/CD pipelines, compliance adherence
Preferred skills
Cloud platforms (AWS/Azure/GCP), Infrastructure as Code (Terraform/Ansible), AI-based solution development, performance bottleneck resolution, eventually consistent system management, PCI-DSS/SOX compliance knowledge
Technologies
Python, Bash, Go, Grafana, ELK stack, Splunk, Docker, Kubernetes, Prometheus, Terraform, Ansible, CloudFormation, AWS, Azure, GCP
Responsibilities
Lead and mentor a globally diverse SRE team, monitor and manage system health and performance, develop automation tools for provisioning and alerting, participate in on-call rotations, collaborate on reliability-focused features, improve observability and logging
Seniority
Mid-level, hands-on IC with people management