Core
Administer production environment changes, lead on-call recovery, and support CI/CD pipelines to ensure operational health and reliability of supported applications.
Role type
Senior DevOps Engineer / Site Reliability Engineer (SRE)
Builds
Production applications and CI/CD pipelines
Domain
Technology Operations / DevOps
Deliverable
production ML models | product features | infrastructure
Required skills
Scripting (Shell, Python, Perl), Linux OS, Networking (TCP/IP, HTTP, Socket, VPN, Firewall, LB), Monitoring (Splunk, Dynatrace), Oracle Database, SQL
Preferred skills
Java, ITSM tools (Remedy, Maximo), Project management, Customer support
Technologies
Shell, Python, Perl, Linux, Splunk, Dynatrace, Oracle, SQL, Java
Responsibilities
Administer production changes and improve operational health; Lead on-call activities for system recovery and post-mortems; Partner with dev teams for smooth releases; Contribute automation and monitoring improvements; Support customer issue diagnosis; Analyze ITSM activities for feedback; Support pre-launch design and capacity planning; Monitor live system availability and latency; Scale systems via automation; Lead DevOps automation best practices; Mentor junior resources.
Seniority
Senior, hands-on IC with mentorship