Staff Software Engineer, Production Engineering
Core
Design, build, and operate the production infrastructure (compute, networking, Kubernetes, workflow orchestration) powering Harvey's AI platform and enterprise services.
Role type
Staff Software Engineer, Production Engineering
Builds
Global compute and network infrastructure, Kubernetes platform, workflow orchestration, and production operations for AI workloads.
Domain
Cloud Infrastructure & AI Platform Engineering
Deliverable
production ML models | infrastructure
Required skills
Large-scale cloud infrastructure (AWS/Azure/GCP), Kubernetes production operations, distributed systems, Infrastructure-as-Code (Terraform/Pulumi), capacity planning, observability, infrastructure security, technical leadership, cross-functional collaboration
Preferred skills
AI/ML or LLM infrastructure at scale, GPU fleet management, multi-cloud environments, internal platform development
Technologies
Kubernetes, Terraform, Pulumi, AWS, Azure, Google Cloud Platform
Responsibilities
Design and operate production infrastructure for AI workloads; drive technical direction for compute, networking, and orchestration; build reusable patterns and tooling for engineering teams; manage on-call rotation and lead incident response; optimize infrastructure cost and performance; establish security foundations and compliance controls.
Seniority
Staff, hands-on IC with strategic influence
