Middle SRE/Infrastructure Automation Engineer
Core
Design, build, and maintain Grafana dashboards and telemetry visualizations to monitor system performance, latency, error rates, saturation, and overall platform health; develop, test, and maintain modular Ansible playbooks to automate infrastructure provisioning, application configuration, patching, and operational workflows.
Role type
Middle SRE/Infrastructure Automation Engineer
Builds
Automated infrastructure deployments, observability solutions, and platform reliability initiatives
Domain
Enterprise IT services and consulting
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Ansible playbooks, Prometheus monitoring and alerting, Grafana dashboards, Git version control, CI/CD workflows, Python or Bash scripting, Linux administration (Ubuntu, RHEL)
Preferred skills
Kubernetes, containerized workloads, metrics storage platforms (Mimir, Thanos), GitLab CI, ITIL processes, change management
Technologies
Ansible, AWX, Prometheus, Grafana, Git, Jira, Python, Bash, Linux, Kubernetes, GitLab CI
Responsibilities
Design and maintain Grafana dashboards for system performance and platform health monitoring; Configure and maintain observability solutions including Prometheus monitoring and alerting; Develop, test, and maintain modular Ansible playbooks for infrastructure automation; Support enterprise automation through AWX by managing centralized execution and reusable workflows; Maintain Infrastructure as Code (IaC) repositories using Git with peer code reviews; Actively participate in Agile ceremonies including Sprint Planning, Daily Stand-ups, and Retrospectives
Seniority
Middle, hands-on IC