Site Reliability Engineer, Tech Lead
Core
Build and maintain Loadsmart's internal platform to ensure application safety, reliability, and operational excellence for a logistics technology company.
Role type
Senior Site Reliability Engineer, Tech Lead
Builds
Internal platform, critical systems, and infrastructure for freight logistics operations
Domain
Logistics technology, Cloud Infrastructure, DevOps
Deliverable
production ML models | infrastructure
Required skills
Cloud computing, SRE/DevOps, Kubernetes, Docker, AWS, CI/CD pipelines, Terraform, Ansible, Linux/UNIX, Python, Bash, monitoring, alerting, incident management, networking, operating systems, system engineering, project management, mentoring
Preferred skills
AI agents, agentic workflows, LLMs, MCP servers/gateways, PostgreSQL, DBA responsibilities
Responsibilities
Design, deploy, and operate critical systems balancing reliability, cost, and agility; drive reliability projects with engineering teams; perform troubleshooting and root-cause analysis; ensure adherence to Service Level Agreements and Objectives; provide infrastructure support during off-hours; take ownership of software infrastructure projects; conduct code and specification reviews.
Seniority
Senior, hands-on IC with leadership responsibilities