Senior Site Reliability Engineer - Undersea Dominance
Core
Build and operate infrastructure for operational and production systems, managing CI/CD pipelines and deploying applications across cloud and edge environments.
Role type
Senior Site Reliability Engineer (MLOps/DevOps)
Builds
CI/CD pipelines, containerized applications, and infrastructure for machine learning models and defense platforms
Domain
Defense technology, maritime autonomous systems, cloud infrastructure
Deliverable
production ML models | infrastructure
Required skills
Python, CI/CD tools (GitHub Actions, Jfrog Artifactory, CircleCI), Infrastructure as Code (Terraform, Ansible), Cloud platforms (Azure, AWS, GCP), Containerization (Docker), Container orchestration (Kubernetes), Logging and monitoring (ELK Stack, Prometheus, Grafana), Parallel computing frameworks (CUDA, OpenCL)
Preferred skills
C++, Rust, Go, Observability concepts, DevOps and MLOps security best practices
Technologies
GitHub Actions, Jfrog Artifactory, Terraform, Ansible, Azure, AWS, GCP, Docker, Kubernetes, ELK Stack, Prometheus, Grafana, MLflow, Kubeflow, CUDA, OpenCL
Responsibilities
Develop and maintain CI/CD pipelines for ML models and applications; Automate infrastructure provisioning and management; Implement containerization and orchestration solutions; Establish monitoring and logging solutions; Collaborate with cross-functional teams to deploy systems
Seniority
Senior, hands-on IC