Senior Software Engineer - Distributed Systems
Core
Design, build, and operate distributed systems for metrics ingestion, processing, and alerting at scale, while developing tooling for AI agents to interact with observability infrastructure.
Role type
Senior IC distributed systems engineer (observability)
Builds
Production-grade metrics pipelines, alerting services, and agentic engineering tooling for AI agents
Domain
Cloud-native infrastructure, observability, distributed systems
Deliverable
production ML models | infrastructure
Required skills
distributed systems fundamentals, cloud-native architecture, Go, Python, Kubernetes, infrastructure as code, API design, fault tolerance, high availability
Preferred skills
agentic engineering (MCP servers, agent-facing APIs), Prometheus-style monitoring, time-series metrics pipelines, GitOps
Technologies
Go, Python, Kubernetes, Helm, Terraform, gRPC, HTTP
Responsibilities
Design and debug distributed systems across cloud-native environments; Develop platform services for metrics ingestion and alerting; Shape AI agent interactions via tooling and APIs; Collaborate with SRE teams to deliver resilient infrastructure
Seniority
Senior, hands-on IC