Site Reliability Engineer, Full-Stack Performance
Core
Build self-service infrastructure, observability stacks, and AI gateway plumbing to enable high-volume feature teams to own their system reliability and performance.
Role type
Senior SRE (Full-Stack Performance & Observability)
Builds
End-to-end observability stack, AI gateway reliability guardrails, automated incident response tooling, and resilient delivery pipelines.
Domain
Real Estate Technology / Distributed Systems / AI Infrastructure
Deliverable
production ML models | infrastructure | product features
Required skills
Java, TypeScript, Kubernetes, APM instrumentation, distributed tracing, CI/CD pipeline optimization, database query profiling, frontend performance profiling (RUM/Core Web Vitals), CLI tooling, system architecture under load.
Preferred skills
Datadog, AWS resource optimization, managing third-party AI service limits, Core Web Vitals expertise.
Responsibilities
Build and instrument end-to-end observability for backend and frontend; operationalize AI gateway with cost and usage tracking; integrate AI-driven insights for automated incident diagnosis; partner on architecting high-traffic resilient paths; optimize delivery pipelines for reliability.
Seniority
Senior, hands-on IC