Custom Software Engineer
Core
Design, build, and configure highly reliable, scalable, and performant systems across the enterprise, acting as the primary point of contact for SRE initiatives.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Production-grade infrastructure, monitoring tooling, and observability standards for enterprise SaaS environments.
Domain
Cloud Infrastructure & Site Reliability Engineering
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Site Reliability Engineering, Infrastructure as Code (IaC), Cloud Platform Management, System Observability, Incident Response, Technical Leadership
Preferred skills
DevOps practices, AWS Administration, Agile methodology
Technologies
Datadog, AWS CloudWatch, Terraform, AWS, Azure, OCI, C#, Go, Python, Java
Responsibilities
Design and optimize environment performance, reliability, and scalability; Implement monitoring tooling for visibility; Develop coding assignments for reliability and alerting; Educate teams on SRE principles; Triage and resolve production issues; Provide technical leadership and mentorship.
Seniority
Senior, hands-on IC with leadership responsibilities