ForgeApply
Try it free

ForgeApply · Job listing

Incident and Problem Management Analyst

Inspire

Atlanta Support Center, USonsite

See all 76 open roles at Inspire

Tailor your resume for this Inspire job in about a minute.

ForgeApply tailors your resume and cover letter to this exact posting, then hands you a ready-to-submit application for Inspire's site. Free trial, no card required.

About this role

The Incident & Problem Management Analyst is responsible for ensuring a cohesive Incident Management, Major Incident Management, and Problem Management practice across Inspire Brands Shared Services. This role leads the Major Incident Management process supporting Shared Services, E-Commerce, Infrastructure, and other critical enterprise platforms.

The analyst serves as the central coordinator during major incidents, leading technical bridge calls, driving rapid engagement of support teams, managing executive communications, and ensuring timely service restoration. This role is responsible for crafting clear, concise, and business-focused communications for senior leadership throughout the incident lifecycle, translating complex technical issues into actionable business updates.

In addition to Major Incident Management responsibilities, the analyst owns key aspects of the Problem Management practice, driving root cause analysis, coordinating corrective actions, and identifying opportunities for long-term incident reduction through ITIL-aligned best practices. Success in this role requires the ability to quickly develop and maintain a strong functional understanding of Inspire Brands' critical technology ecosystem, including MDBP, IDP, digital commerce platforms, integrations, APIs, infrastructure components, and end-to-end transaction flows. The analyst must be capable of rapidly assessing business impact, understanding system dependencies, and coordinating technical teams during high-severity incidents. Responsibilities Major Incident Management • Lead and coordinate Major Incident Management activities for enterprise technology services and platforms. • Serve as Incident Commander during high-priority incidents and service disruptions. • Facilitate and lead technical bridge calls during critical incidents. • Engage and coordinate appropriate technical resources, vendors, and business stakeholders to expedite resolution. • Drive accountability across support teams and ensure clear ownership of resolution activities. • Maintain command and control throughout the incident lifecycle. • Quickly assess business impact and communicate risk, customer impact, and service degradation to stakeholders. • Develop and distribute executive-facing communications, including: • Initial incident notifications • Status updates • Business impact assessments • Executive summaries • Resolution communications • Post-incident reports

• Translate complex technical information into concise, business-focused messaging for senior leadership. • Coordinate post-incident reviews and ensure action items are identified and tracked to completion. • Support high-visibility business events, hypercare periods, and executive escalations when service disruptions occur.

Problem Management • Identify, classify, and manage problem records to prevent recurring incidents. • Analyze incident trends and recurring issues to identify opportunities for systemic improvements. • Drive Root Cause Analysis (RCA) efforts following major incidents and significant service disruptions. • Coordinate internal and external stakeholders to investigate root causes. • Ensure corrective actions are assigned, tracked, and completed. • Manage problem records through closure while ensuring proper documentation and knowledge transfer. • Promote proactive problem management practices across technology teams. • Drive continuous service improvement initiatives that reduce operational risk and improve service stability.

Operational Excellence & Service Governance • Coordinate resources and processes required to resolve enterprise-wide outages and service disruptions. • Support the continual improvement of Incident, Major Incident, and Problem Management processes. • Maintain and enhance operational documentation, standard operating procedures, and knowledge articles. • Ensure quality procedures and solutions are documented and readily available to support teams. • Collaborate with technical and business teams to improve operational maturity and service reliability. • Support reporting, service governance, and operational metrics initiatives. • Contribute to ITSM maturity improvements aligned to ITIL best practices and organizational objectives.

Technical Understanding & Impact Assessment • Develop and maintain a high-level understanding of Inspire Brands' technology ecosystem, including MDBP, IDP, digital platforms, integrations, APIs, and supporting infrastructure. • Understand application dependencies, transaction flow paths, integration points, and service relationships. • Quickly identify and assess outage impacts across interconnected business systems. • Evaluate risks and potential customer impact during service disruptions. • Partner with application, infrastructure, network, security, and vendor teams to accelerate incident diagnosis and resolution. • Support continual refinement of service dependency mapping and operational readiness processes.

Communication & Stakeholder Management • Communicate effectively with business stakeholders, technical teams, vendors, management, and senior leadership. • Develop and deliver executive-level incident communications during high-pressure situations. • Facilitate collaboration across multiple technical teams and organizational functions. • Influence stakeholders and drive accountability toward resolution and corrective action completion. • Maintain professionalism and composure during critical incidents and executive escalations.

Qualifications

Education • Bachelor's Degree in Computer Science, Information Technology, Management Information Systems, or a related field.

Experience • Minimum of 3 years of experience in a medium to large enterprise environment. • 3+ years of experience defining, implementing, and improving IT Service Management (ITSM) processes. • 3+ years of experience supporting Incident Management, Major Incident Management, and Problem Management practices. • Experience leading inci

Salary insight

This posting doesn't disclose pay. Across 650 Atlanta jobs with disclosed salaries on ForgeApply, the median is $135k.

Based on live postings with disclosed pay on ForgeApply; refreshed daily. Not an estimate of this employer's offer.

Tailor your resume for this Inspire role before you apply.

Tailor my resume for this job

Similar jobs

Free ATS checker · No Salary on the Job Posting? How to Find the Number Before You Interview