SRE Software Systems Engineer IV
Core
Drive platform reliability, auto-scaling, and cloud cost efficiency for high-throughput data pipelines and AI platforms powering travel retailing.
Role type
Senior IC Site Reliability Engineer (Data Intelligence & AI Operations)
Builds
High-throughput data pipelines, AI/ML platforms, and Google Cloud Platform infrastructure for travel partners.
Domain
Travel technology, Cloud Infrastructure, Machine Learning Operations
Deliverable
production ML models | infrastructure
Required skills
Google Cloud Platform, Terraform, Python, SQL, Apache Airflow, BigQuery, Kubernetes, Generative AI, Infrastructure as Code
Preferred skills
Incident response, schema migrations, observability, automation
Technologies
Google Cloud Platform, Terraform, BigQuery, Apache Airflow, Python, Bash, Kubernetes
Responsibilities
Modularize and manage core GCP services using Infrastructure as Code; Optimize resource utilization and auto-scaling for ML workloads; Operationalize ML models and RAG architectures via CI pipelines; Build automated failovers and schema migrations; Use generative AI to streamline observability and incident response.
Seniority
Senior, hands-on IC