Prodops Engineer 3
Core
Stabilize large-scale production systems and manage critical incidents in a 24/7 environment for customer-facing software.
Role type
Senior Production Operations Engineer (SRE)
Builds
Secure, high-quality software via SAST, SCA, and DAST solutions
Domain
Application Security / DevSecOps
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Incident management, Root cause analysis, Automation scripting, Kubernetes orchestration, Cloud platform management, CI/CD pipeline maintenance, Technical mentorship, Runbook standardization
Preferred skills
Distributed systems architecture, Disaster recovery planning, Shift leadership
Technologies
Go, Python, Shell, Perl, Docker, Kubernetes, Helm, Terraform, Jenkins, Harness, GitHub Actions, ArgoCD, GitLab CI, Prometheus, Grafana, ELK, Datadog, New Relic, Loki, Git, GitHub, GitLab
Responsibilities
Own and manage critical production incidents end-to-end; Perform root cause analysis and drive corrective actions; Automate operational tasks to reduce toil; Maintain dashboards, alerts, runbooks, and SOPs; Handle customer-facing communications during incidents; Guide junior engineers and support shift handovers; Act as a technical reviewer for reliability-critical changes; Influence architecture decisions with operability and reliability in mind.
Seniority
Senior, hands-on IC with mentorship responsibilities