Senior Site Reliability Engineer, Booking Services Search Group - Search Department (SED)
Core
Ensuring reliability of large-scale distributed Search platforms for Rakuten's e-commerce and ecosystem services.
Role type
Senior Site Reliability Engineer
Builds
Production environments for new clients, search services for Rakuten Travel and Leisure businesses
Domain
E-commerce, Travel, and Internet Ecosystems
Deliverable
infrastructure
Required skills
Linux system performance tuning, distributed systems performance analysis, configuration management (Chef, Ansible), observability (Prometheus, Grafana, Loki), container orchestration (Kubernetes), automation scripting (Shell, Python), computer networking protocols, Java service management
Preferred skills
CI/CD automation (Jenkins), big data/streaming technologies (Spark, Solr, Cassandra, Kafka), cloud storage (Ceph, S3, MinIO), Git workflows, software development
Technologies
Kubernetes, Chef, Ansible, Prometheus, Grafana, Loki, Jenkins, Spark, Solr, Cassandra, Kafka, Ceph, S3, MinIO, Java, Shell, Python
Responsibilities
Design and deploy production environments from hardware to service level, monitor production systems and handle on-call duties, troubleshoot and resolve incidents, manage production releases including nighttime operations, drive tooling and automation stack improvements, collaborate on testing plans including performance and security drills
Seniority
Senior, hands-on IC