Manager Technology
Core
Managing a team of 6-8 engineers to design, build, and operate scalable cloud-native platform solutions, focusing on reliability, operability, and developer experience.
Role type
Senior IC Platform Engineering Manager / SRE Lead
Builds
Internal developer platforms, self-service tooling, and production-grade infrastructure services for engineering teams.
Domain
Cloud-native infrastructure, Platform Engineering, Site Reliability Engineering (SRE)
Deliverable
production ML models | product features | infrastructure
Required skills
Kubernetes (AKS), Terraform, GitOps (ArgoCD, Helm/Kustomize), SRE practices (SLOs, SLIs, error budgets), incident management, capacity planning, Spec Driven Development (SDD), AI tool integration for infrastructure automation
Preferred skills
Observability tooling (Datadog, Prometheus, Grafana), microservices architecture, service mesh, cloud-native patterns, AI-assisted development workflows
Technologies
AKS, Terraform, ArgoCD, Argo Workflows, Helm, Kustomize, GitHub Actions, Azure DevOps, Datadog, Prometheus, Grafana, Co-pilot, Claude
Responsibilities
Manage day-2 operations for AKS clusters including upgrades, patching, and capacity planning; Establish SRE practices including monitoring, alerting, and performance tuning; Design and implement scalable platform solutions with architects and product owners; Mentor team members and drive adoption of platform capabilities and AI tools; Define infrastructure contracts and APIs using Spec Driven Development; Conduct Production Readiness Reviews and manage incident response processes.
Seniority
Senior, hands-on IC with team management