SonicJobs Logo
Left arrow iconBack to search

Major Incident Manager

ZENITH INFOTEK LLC
Posted 6 days ago, valid for 21 days
Location

Brush Prairie, WA, US

Salary

$110,000 - $120,000 per year

Contract type

Full Time

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.

Sonic Summary

info
  • The job involves managing crisis commands and orchestrating incidents, requiring a focus on service restoration during high-impact events.
  • Candidates should have experience in leading Major Incident Command Bridges and coordinating efforts among various teams.
  • Effective communication with stakeholders and executive leadership is crucial, providing timely updates and managing escalations as needed.
  • The role includes governance responsibilities such as facilitating post-incident reviews and ensuring smooth transitions to Problem Management.
  • A minimum of 5 years of experience is required, with a salary range of $90,000 to $120,000, and benefits including 401(k), dental, and vision insurance.
Benefits:
  • 401(k)
  • Dental insurance
  • Vision insurance


Job Summary
 
1. Crisis Command & Incident Orchestration
  • ​Trigger & Mobilization: Instantly invoke the Major Incident Management process upon notification of a Priority 1 (P1) or high-impact Priority 2 (P2) event.
  • ​Bridge Management: Chair and lead the Major Incident Command Bridge / War Room. Coordinate internal engineering, infrastructure, application, and 3rd-party vendor teams.
  • ​Drive Restoration: Maintain absolute focus on service restoration and workarounds rather than immediate root cause analysis.
​2. Stakeholder & Executive Communication
  • ​Broadcast timely, clear, and non-jargon status updates to executive leadership, business unit heads, and key stakeholders at fixed cadences (e.g., every 15–30 minutes).
  • ​Manage escalation paths to engage senior technical leads or external suppliers when resolution stalls.
​3. Governance & Post-Incident Management (PIR)
  • ​Hand-off to Problem Management: Ensure seamless transition of the incident to the Problem Management team for Root Cause Analysis (RCA) once service is restored.
  • ​Post-Incident Reviews (PIR): Facilitate PIR sessions to capture timelines, evaluate response effectiveness, and document lessons learned.
  • ​Emergency Changes: Authorize and log Emergency Change Requests (ECRs) required for immediate fixes in compliance with ITIL Change Management.
​4. Process & KPI Reporting
  • ​Track key performance indicators, including Mean Time to Detect (MTTD), Mean Time to Restore Service (MTRS), and SLA compliance.
  • ​Continuously refine Major Incident playbooks, escalation matrices, and response workflows.



Learn more about this Employer on their Career Site

Apply now in a few quick clicks

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.