- Location
- Bangalore, India
- Type
- Full-time
- Department
- Engineering
- Seniority
- Lead
- Source
- Workday
Description
Company Overview
At Motorola Solutions, we believe that everything starts with our people. We’re a global close-knit community, united by the relentless pursuit to help keep people safer everywhere. We build and connect technologies to help protect people, property and places. Our solutions foster the collaboration that’s critical for safer communities, safer schools, safer hospitals, safer businesses, and ultimately, safer nations. Connect with a career that matters, and help us build a safer future.
Department Overview
The Design & Tools (D&T) team is a strategic component of Centralized Managed Support Operations (CMSO), responsible for providing exceptional Service Design & innovative technology solutions to enable and empower MSI Centralized Managed Support Operations to meet and exceed customer's expectations.
Job Description
Position Overview
We are seeking a highly skilled and strategic Principal Observability & Event Management Engineer to champion enterprise-wide initiatives that improve monitoring effectiveness, reduce alert noise, and accelerate incident response.
In this dual-impact role, you will act as both a strategic leader and a hands-on technical expert. You will define the enterprise standards for multi-domain monitoring (infrastructure, cloud, application, network, batch, and mainframe) while actively designing and configuring our Event Management platforms—specifically Oracle Unified Assurance (Assure1) and IBM Netcool. The ideal candidate will leverage data insights, machine learning, and automation to transform massive "event storms" into high-fidelity, actionable signals.
Key Responsibilities
Strategic Leadership & Governance
Drive Enterprise Initiatives: Own and lead cross-functional, enterprise-wide initiatives focused on improving monitoring effectiveness, optimizing alert quality, and reducing incident volumes.
Define Standards: Establish corporate standards, approaches, and best practices for alert optimization across infrastructure, application, batch, network, and mainframe environments.
Culture & Mentorship: Mentor, guide, and evangelize a strong observability and optimization mindset across engineering and operations teams.
Continuous Improvement: Identify systemic gaps in monitoring design and lead long-term, architectural improvements to eliminate redundant or low-value alerts.
Platform Engineering & Administration
Infrastructure Design: Implement, maintain, and scale Oracle Assure1 / Unified Assurance servers and architectures for high-availability (HA) enterprise environments.
Event Correlation & Enrichment: Develop advanced Object-Server configurations, triggers, and correlation rules to distinguish isolated anomalies from critical root causes.
Multi-Vendor Integration: Integrate multi-vendor telemetry, logs, faults, and data from disparate monitoring domains into a single, unified Event Management UI.
Operations & Backend Support: Manage backend databases (e.g., MySQL, OpenSearch), maintain platform availability, and troubleshoot complex backend issues.
Automation, AIOps & Incident Response
Incident Automation: Write robust automation scripts and workflows to bridge monitoring tools with ITSM platforms (such as ServiceNow) to accelerate incident response times.
AIOps & Root Cause Analysis (RCA): Apply Machine Learning (ML) algorithms, topological discovery, and data-driven approaches to enhance event correlation, suppress noise, and automate self-healing or ticket generation.
Basic Requirements
Required Qualifications & Skills
Experience & Education
Total Experience: 5+ years of dedicated experience in enterprise Event Management, Fault Management, Service Level Management, or production environments.
Initiative Leadership: 3+ years of proven experience leading large-scale, cross-functional improvements in monitoring effectiveness, alert optimization, or incident reduction.
Technical Expertise
Platform Expertise: Hands-on, deep technical experience with Oracle Unified Assurance / Assure1 and/or IBM Netcool suites (OMNIbus, Netcool Impact, ITM).
OS & Scripting: Strong proficiency in Linux Administration (RHEL) and expert-level scripting skills (Python, Bash, Perl, or Shell) for building event parsers and API integrations.
Databases & APIs: Solid understanding of backend database management, SQL queries, REST APIs, and tool customization using rule files and OpenSearch.
Domain Knowledge: Strong exposure to multi-domain infrastructure (Cloud, Network, Telecommunications, Mainframe) paired with a solid understanding of ITIL Event Management principles.
Travel Requirements
None
Relocation Provided
None
Position Type
ExperiencedReferral Payment Plan
NoEEO Statement
Motorola Solutions is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion or belief, sex, sexual orientation, gender identity, national origin, disability, veteran status or any other legally-protected characteristic.
We are proud of our people-first and community-focused culture, empowering every Motorolan to be their most authentic self and to do their best work to deliver on the promise of a safer world. If you’d like to join our team but feel that you don’t quite meet all of the preferred skills, we’d still love to hear why you think you’d be a great addition to our team.