Principal Observability & Event Management Engineer
Core
Lead enterprise-wide observability initiatives to improve monitoring effectiveness, reduce alert noise, and accelerate incident response across multi-domain environments.
Role type
Principal Observability & Event Management Engineer (Strategic IC)
Builds
Enterprise Event Management platforms (Oracle Assure1, IBM Netcool) and automated incident response workflows
Domain
IT Operations, Observability, AIOps, Telecommunications Infrastructure
Required skills
Oracle Unified Assurance (Assure1), IBM Netcool, Linux Administration (RHEL), Python, Bash, Perl, SQL, REST APIs, Machine Learning, Event Correlation, Incident Automation
Preferred skills
Mainframe monitoring, Topological discovery, ITIL Event Management principles
Technologies
Oracle Assure1, IBM Netcool, MySQL, OpenSearch, ServiceNow
Responsibilities
Define corporate standards for alert optimization across infrastructure, application, batch, network, and mainframe environments; Design and configure Event Management platforms; Develop advanced correlation rules to distinguish anomalies from root causes; Write automation scripts to bridge monitoring tools with ITSM platforms; Apply ML algorithms to enhance event correlation and automate self-healing; Mentor engineering and operations teams on observability best practices
Seniority
Principal, strategic leadership with hands-on technical execution (via careerplan.io/jobs/R66422-principal-observability-event-management-engineer-at-motorolasolutions)