Principal Observability & Event Management Engineer
Motorola Solutions
- Location
- Bangalore, India
- Work model
- On-Site
- Level
- Principal
- H-1B history
- 64 approvals (FY2023)
- Posted
- Aug 25, 2026
Skills
About this role
Company Overview At Motorola Solutions, we believe that everything starts with our people. We’re a global close-knit community, united by the relentless pursuit to help keep people safer everywhere. We build and connect technologies to help protect people, property and places. Our solutions foster the collaboration that’s critical for safer communities, safer schools, safer hospitals, safer businesses, and ultimately, safer nations. Connect with a career that matters, and help us build a safer future. Department Overview The Design & Tools (D&T) team is a strategic component of Centralized Managed Support Operations (CMSO), responsible for providing exceptional Service Design & innovative technology solutions to enable and empower MSI Centralized Managed Support Operations to meet and exceed customer's expectations.
Job Description
Position Overview We are seeking a highly skilled and strategic Principal Observability & Event Management Engineer to champion enterprise-wide initiatives that improve monitoring effectiveness, reduce alert noise, and accelerate incident response. In this dual-impact role, you will act as both a strategic leader and a hands-on technical expert. You will define the enterprise standards for multi-domain monitoring (infrastructure, cloud, application, network, batch, and mainframe) while actively designing and configuring our Event Management platforms—specifically Oracle Unified Assurance (Assure1) and IBM Netcool . The ideal candidate will leverage data insights, machine learning, and automation to transform massive "event storms" into high-fidelity, actionable signals.
Key Responsibilities
Strategic Leadership & Governance Drive Enterprise Initiatives: Own and lead cross-functional, enterprise-wide initiatives focused on improving monitoring effectiveness, optimizing alert quality, and reducing incident volumes. Define Standards: Establish corporate standards, approaches, and best practices for alert optimization across infrastructure, application, batch, network, and mainframe environments. Culture & Mentorship: Mentor, guide, and evangelize a strong observability and optimization mindset across engineering and operations teams. Continuous Improvement: Identify systemic gaps in monitoring design and lead long-term, architectural improvements to eliminate redundant or low-value alerts. Platform Engineering & Administration Infrastructure Design: Implement, maintain, and scale Oracle Assure1 / Unified Assurance servers and architectures for high-availability (HA) enterprise environments. Event Correlation & Enrichment: Develop advanced Object-Server configurations, triggers, and correlation rules to distinguish isolated anomalies from critical root causes. Multi-Vendor Integration: Integrate multi-vendor telemetry, logs, faults, and data from disparate monitoring domains into a single, unified Event Management UI. Operations & Backend Support: Manage backend databases (e.g., MySQL, OpenSearch), maintain platform availability, and troubleshoot complex backend issues. Automation, AIOps & Incident Response Incident Automation: Write robust automation scripts and workflows to bridge monitoring tools with ITSM platforms (such as ServiceNow ) to accelerate incident response times. AIOps & Root Cause Analysis (RCA): Apply Machine Learning (ML) algorithms, topological discovery, and data-driven approaches to enhance event correlation, suppress noise, and automate self-healing or ticket generation.
Basic Requirements
Required Qualifications & Skills Experience & Education Total Experience: 5+ years of dedicated experience in enterprise Event Management, Fault Management, Service Level Management, or production environments. Initiative Leadership: 3+ years of proven experience leading large-scale, cross-functional improvements in monitoring effectiveness, alert optimization, or incident reduction. Technical Expertise Platform Expertise: Hands-on, deep technical experience with Oracle