Technical Product Owner — Observability Event Management
Core
Own the functional design and capability roadmap for an enterprise observability event management platform that converts raw telemetry and alerts into actionable incidents, automated remediation, and AI-assisted root cause analysis.
Role type
Senior Technical Product Owner (Observability & Event Management)
Builds
A vendor-agnostic, composable event management platform with event correlation, anomaly detection, and agentic self-healing capabilities.
Domain
Enterprise IT Operations, Observability, SRE, AI/ML in Infrastructure
Deliverable
production ML models | product features
Required skills
Functional design of complex technical systems, roadmap ownership, deep understanding of distributed systems and event-driven architectures, defining rigorous product requirements and acceptance criteria, guiding engineering design trade-offs, translating SRE needs into product capabilities.
Preferred skills
Hands-on experience with observability tools (Datadog, Dynatrace, Splunk, ServiceNow, PagerDuty, Grafana, Prometheus, OpenTelemetry), SRE practices, event correlation logic, anomaly detection, remediation orchestration, GenAI/agentic-AI product design, ITIL 4 frameworks.
Technologies
Datadog, Dynatrace, Splunk ITSI, ServiceNow ITOM, PagerDuty, BigPanda, Moogsoft, New Relic, Grafana, Prometheus, OpenTelemetry, GenAI, MCP, Retrieval-Augmented Generation
Responsibilities
Define end-to-end functional design for event ingestion, correlation, enrichment, and remediation; translate product strategy into a capability roadmap; lead technical discovery with SREs and architects; define functional specifications and user stories; partner with engineers on API and data model design; shape agentic capabilities and human-governed workflows; validate functional quality and measurable value delivery.
Seniority
Senior, hands-on IC