Director -Software Engineering, Infrastructure
Core
Lead a high-impact engineering organization focused on SQL reliability, data platform stability, and observability to improve uptime, customer experience, and cost efficiency.
Role type
Director of Infrastructure Software Engineering (SQL Reliability & Observability)
Builds
Unified observability platform, proactive monitoring systems, and AI-driven anomaly detection for distributed databases.
Domain
Cloud Infrastructure, Database Reliability, Observability, SRE
Deliverable
production ML models | infrastructure
Required skills
SQL and relational database systems, cloud platforms (Azure), observability platforms (metrics, logs, tracing), Kubernetes, performance engineering, distributed systems architecture, engineering leadership, team hiring and development, cross-functional alignment.
Preferred skills
NoSQL systems (Cosmos DB, MongoDB, Redis, Kafka), OpenTelemetry, CI/CD reliability practices, AI/ML-driven monitoring or root cause analysis systems.
Technologies
Azure SQL, PostgreSQL, MySQL, Cosmos DB, MongoDB, Redis, Kafka, OpenTelemetry, Kubernetes, Azure.
Responsibilities
Define SLOs/SLIs and reliability frameworks; oversee performance, scaling, and availability of database systems; lead development of a unified observability platform; implement synthetic monitoring and AI-driven anomaly detection; hire and mentor senior engineers; partner with Infrastructure, Product, and Finance teams.
Seniority
Director, hands-on IC with leadership responsibilities
