Staff DevOps Engineer
Core
Architect, develop, and scale core services, real-time data pipelines, and distributed infrastructure to power cutting-edge AI products and bridge the gap between complex AI models and production-ready systems.
Role type
Staff DevOps Engineer (Distributed Systems & AI Infrastructure)
Builds
Highly available backend systems, distributed data pipelines, scalable cloud infrastructure, and real-time AI applications.
Domain
Aerospace / AI Infrastructure / Distributed Systems
Deliverable
production ML models | infrastructure
Required skills
Java, Python, Rust, Apache Flink, Kubernetes, Docker, Terraform/OpenTofu, Ansible, GitOps (Argo CD), PostgreSQL, Redis/Valkey, observability stacks (Grafana, Prometheus, ELK)
Preferred skills
Apache Pulsar/Kafka, Rust async runtimes (Tokio), WebRTC/LiveKit, audio codecs (Opus), MLOps frameworks, GPU orchestration, high-availability real-time systems
Technologies
Apache Flink, Kubernetes, Docker, Terraform, OpenTofu, Ansible, Argo CD, Grafana, Prometheus, ELK, PostgreSQL, Redis, Valkey, AWS, GCP, Azure, Rust, Java, Python
Responsibilities
Architect and maintain highly available backend systems and scalable cloud infrastructure; build and manage stream-processing jobs for high-frequency data streams; containerize, deploy, and orchestrate scalable ML models; own the full development lifecycle from design to capacity planning; maximize system performance and reliability through advanced debugging and root cause analysis; mentor team members on DevOps practices and distributed systems best practices.
Seniority
Staff, hands-on IC with mentorship responsibilities