Spécialiste Fiabilité outils et infrastructures - TGQF
Core
Ensure the viability, stability, and performance of tools and infrastructures supporting studio activities, acting as a technical resource for incident prevention and resolution.
Role type
Senior IC DevOps/Infrastructure Reliability Engineer
Builds
Stable and reliable development environments, automated deployment pipelines, and incident management processes
Domain
Video game development, Cloud infrastructure, Observability
Deliverable
production ML models | infrastructure
Required skills
Software development, System administration, Database administration, Infrastructure automation (Cloud & On-prem), Programming languages, Observability technologies, CI/CD processes, Cloud services, Network infrastructure, Configuration management, Linux/Windows administration, Redundant and scalable architecture design, Code optimization, Task automation
Preferred skills
Agile methodologies, Technical teaching and mentoring, Proactive problem solving, Rapid adaptation to fast-paced environments
Technologies
Grafana, Splunk, Elasticsearch, Prometheus, OpenTelemetry, Docker, Git, Terraform, Ansible, Chef
Responsibilities
Accompany development teams in technology choices to improve system visibility and robustness; Automate processes to facilitate team work; Implement tools and methods for secure service deployment and incident management; Diagnose and permanently fix anomalies and infrastructure failures; Coordinate resources to restore service level objectives; Design, deploy, secure, and maintain reliable environments; Provide continuous technical support and proactively address issues; Create and maintain deployment guides and technical documentation.
Seniority
Senior, hands-on IC