Principal AI Platform Operations Engineer
Core
Setting long-term AI operations and platform vision, strategy, and governance for enterprise-scale production systems.
Role type
Principal AI Platform Operations Engineer (Strategy & Architecture)
Builds
Enterprise AI/ML deployment platforms, reliability frameworks, and governance standards
Domain
Higher Education / Enterprise AI Operations
Deliverable
production ML models | infrastructure | dashboards & analysis
Required skills
AI operations strategy, reference architecture design, LLMOps/MLOps at scale, AI safety and governance, executive advisory, technical hiring, cross-functional leadership, thought leadership, cloud platform expertise, observability, reliability engineering
Preferred skills
Databricks platform expertise, AWS cloud expertise, infrastructure-as-code, container orchestration
Technologies
Databricks, AWS, Terraform, CloudFormation
Responsibilities
Define multi-year AI operations technical vision and strategy; Establish foundational reference architectures and engineering principles; Lead evaluation and adoption of frontier AI operations paradigms; Drive complex AI operations programs including large-scale platform and migration initiatives; Advise executives on AI operations strategy, risk, and investment priorities; Define AI operations governance including responsible AI standards and compliance frameworks; Incubate emerging operational capabilities and sponsor proof-of-concept initiatives; Serve as primary external technical spokesperson; Design and evolve team operating model and engineering culture; Mentor Staff, Senior, and II-level engineers; Author internal and external thought leadership; Partner with legal, compliance, and privacy teams on regulatory requirements
Seniority
Principal, hands-on IC with strategy & mentorship
