Senior Automation Engineer - Dynatrace
Core
Design and implement automation scripts, workflows, and integrations using Dynatrace telemetry to reduce operational toil and enable AIOps-driven investigation and remediation.
Role type
Senior IC automation engineer (AIOps/Dynatrace)
Builds
Automation scripts, workflows, runbooks, and integrations between Dynatrace and client platforms (ServiceNow, cloud, CI/CD)
Domain
IT Operations / SRE / AIOps / Observability
Deliverable
production ML models | product features | dashboards & analysis | research | client delivery | infrastructure | physical/clinical work
Required skills
Python, PowerShell, Bash, REST APIs, ServiceNow, Dynatrace APIs, cloud automation, Kubernetes, Terraform, Ansible, Jenkins, GitHub Actions, Splunk, CloudWatch, AI/GenAI, LLM, RAG, AI-agent based IT operations
Preferred skills
Dynatrace Workflows, Dynatrace Intelligence / Davis AI, AWS/Azure, SRE practices (SLI/SLO, error budgets), ROI/toil-reduction calculations
Technologies
Dynatrace, ServiceNow, AWS, Azure, Kubernetes, EKS, Terraform, Ansible, Jenkins, GitHub Actions, Harness, Splunk, CloudWatch, Slack, Microsoft Teams
Responsibilities
Analyze historical tickets and operational data to identify high-volume, repetitive operational toil; Assess automation opportunities based on frequency, complexity, repeatability, risk and business impact; Build an automation opportunity backlog and quantify potential effort/toil reduction; Identify whether each use case is best addressed through rules, scripts, workflow automation, AI-assisted recommendations or agentic automation; Use Dynatrace telemetry, events, problems and service dependencies to identify opportunities for AIOps-driven investigation and remediation; Develop automation scripts and workflows using Python, PowerShell, Bash, REST APIs and existing client tools; Build integrations between Dynatrace and existing client platforms such as ServiceNow, cloud platforms, CI/CD, monitoring and collaboration tools; Design and implement human-gated remediation for production operations; Implement recovery validation covering technical health, service/SLO health and business transaction validation; Develop reusable automation components and runbooks rather than one-off scripts; Support AIOps use cases such as event correlation, incident enrichment, probable-cause analysis, recommendation and controlled remediation; Measure automation outcomes including toil reduction, automation coverage, MTTR improvement and recurring-incident reduction; Work closely with SRE/AIOps architects and operations teams to convert identified opportunities into PDP2 Thin MVP use cases and production-ready automation
Seniority
Senior, hands-on IC