Digital Technology Senior Specialist – Observability & AI Ops
Core
Implement, manage, and optimize enterprise observability platforms (Elastic Stack) to detect issues, troubleshoot faster, and improve service performance across cloud, infrastructure, and application environments.
Role type
Senior IC AIOps & Observability Engineer
Builds
Enterprise observability platforms, dashboards, alerts, and operational reports for technology teams
Domain
Energy technology, Cloud Infrastructure, SRE
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Elastic Stack administration (Kibana, Logstash, Beats, APM, Fleet), Cloud operations (AWS/Azure), Automation & scripting (Python, Bash, PowerShell), Container orchestration (Kubernetes, Docker), Incident management & Root cause analysis, Infrastructure as Code (Terraform, Ansible), Data visualization & dashboarding
Preferred skills
AWS/Azure Associate certification, Elastic Certified Engineer, Open Telemetry, Service meshes, Hybrid deployment experience, Agile/ITIL process knowledge
Technologies
Elastic Stack (Elasticsearch, Kibana, Logstash, Beats, APM, Fleet), AWS, Azure, Kubernetes, Docker, Prometheus, Grafana, Terraform, Ansible, GitHub Actions, Jenkins, Python, Bash, PowerShell, YAML
Responsibilities
Implement and manage the lifecycle of enterprise observability platforms; Onboard 20,000+ devices, infrastructure, cloud services, and applications; Build and maintain dashboards, visualizations, and alerts; Configure log, metric, trace, and APM data collection; Support incident investigation and root cause analysis; Contribute to automation initiatives using scripting and IaC; Participate in platform upgrades, patching, and performance tuning; Create and maintain operational runbooks and documentation.
Seniority
Senior, hands-on IC