Sr. Manager, Monitoring & Observability
Core
Design, engineer, and implement robust global IT monitoring and observability capabilities across on-prem, cloud, and hybrid environments, focusing on automated remediation, proactive alerting, and predictive service impact analysis.
Role type
Sr. Manager, Enterprise Monitoring & Observability
Builds
Enterprise monitoring solutions, automated remediation systems, and event correlation platforms for PVH IT landscape
Domain
Enterprise IT Operations, Cloud & On-Prem Infrastructure, Observability
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Linux and Windows Server administration, Python/JavaScript/PowerShell scripting, REST API and JSON, SQL query writing, architecture design, project management
Preferred skills
Experience with SCOM, APM tools, log management tools, Manager of Managers (MoM), service management standards
Technologies
SCOM, APM tools, log management tools, Manager of Managers (MoM), Linux, Windows Server, Python, JavaScript, PowerShell, SQL
Responsibilities
Plan and implement cost-effective monitoring solutions across production and pre-production environments; Generate and improve alerting with a focus on proactive alerting and self-healing; Build service management and event correlation to transition from isolated alerts to predictive service impact analysis; Partner with operations teams to design integrations among processes and tools; Develop and maintain technical documentation for monitoring configuration and processes; Conduct research on monitoring products and standards to stay abreast of developments
Seniority
Senior, hands-on IC with management responsibilities