Padmi

Production Support Analyst (Command Center) (Pune)

MumbaiPosted 2 months ago
IT supportJuniorFull Time; Regular
Apply at IntraEdge

Opens the source posting on shine.com

Source description

About the role

View original

L1/L2 Monitoring & Incident Management Specialist (Production Support Analyst (Command Center)) Role Overview We are looking for an experienced L1/L2 Monitoring & Incident Management Specialist to support centralized monitoring and incident command operations for business-critical, client-facing applications across both mainframe and distributed environments. The role requires proactive monitoring, rapid incident response, effective stakeholder communication, and coordination across multiple technical teams to ensure service availability and operational excellence. Key Responsibilities - Provide 24x7 monitoring and operational support for client-facing applications and services. - Monitor real-time system alerts and respond promptly to incidents impacting business operations. - Act as the primary point of contact during incidents, ensuring timely communication with clients and stakeholders. - Perform initial incident triage, impact assessment, and preliminary root cause analysis. - Coordinate with Infrastructure, Application, and Vendor teams to facilitate incident resolution and service restoration. - Manage and track incidents through their complete lifecycle, ensuring adherence to defined SLAs. - Handle production issues including batch failures, system alerts, service degradation, and application outages. - Escalate critical issues appropriately through ITSM platforms such as ServiceNow. - Maintain accurate incident records, status updates, and post-incident documentation. Required Skills & Experience - Solid experience with enterprise monitoring and alert management tools. - Hands-on experience with IT Service Management (ITSM) platforms, preferably ServiceNow. - Solid understanding of incident management processes and incident lifecycle management. - Experience working in production support environments with 24x7 operational coverage. - Exposure to mainframe systems, distributed applications, APIs, and integrated technology environments. - Knowledge of service restoration processes, escalation management, and operational support best practices. Key Competencies - Ability to work effectively in high-pressure, mission-critical environments. - Strong analytical and problem-solving skills with the ability to make real-time decisions. - Excellent communication and stakeholder management skills. - Ability to coordinate across multiple teams and drive incidents to resolution. - Strong sense of ownership, accountability, and customer focus. - Ability to prioritize and manage multiple incidents simultaneously while maintaining service quality. Preferred Qualifications - Experience supporting enterprise-scale production environments. - Familiarity with SLA-driven support models and operational governance. - Understanding of ITIL principles and service management frameworks. - Experience working with both mainframe and distributed technology stacks L1/L2 Monitoring & Incident Management Specialist (Production Support Analyst (Command Center)) Role Overview We are looking for an experienced L1/L2 Monitoring & Incident Management Specialist to support centralized monitoring and incident command operations for business-critical, client-facing applications across both mainframe and distributed environments. The role requires proactive monitoring, rapid incident response, effective stakeholder communication, and coordination across multiple technical teams to ensure service availability and operational excellence. Key Responsibilities - Provide 24x7 monitoring and operational support for client-facing applications and services. - Monitor real-time system alerts and respond promptly to incidents impacting business operations. - Act as the primary point of contact during incidents, ensuring timely communication with clients and stakeholders. - Perform initial incident triage, impact assessment, and preliminary root cause analysis. - Coordinate with Infrastructure, Application, and Vendor teams to facilitate incident resolution and service restoration. - Manage and track incidents through their complete lifecycle, ensuring adherence to defined SLAs. - Handle production issues including batch failures, system alerts, service degradation, and application outages. - Escalate critical issues appropriately through ITSM platforms such as ServiceNow. - Maintain accurate incident records, status updates, and post-incident documentation. Required Skills & Experience - Solid experience with enterprise monitoring and alert management tools. - Hands-on experience with IT Service Management (ITSM) platforms, preferably ServiceNow. - Solid understanding of incident management processes and incident lifecycle management. - Experience working in production support environments with 24x7 operational coverage. - Exposure to mainframe systems, distributed applications, APIs, and integrated technology environments. - Knowledge of service restoration processes, escalation management, and operational support best practices.

One address, no account. We’ll tell you when matching roles go live.

More at IntraEdge

Related open roles

View all roles