Incident Manager

🏢 ZainTECH
📍 New Cairo City, EgyptFull-timeOn-site
📅 Posted: 6mo ago🔄 Updated: 6mo ago
CV%
✨ AI Summary
The Incident Manager oversees end-to-end incident management for cloud and infrastructure services, including AWS, Azure, and OCI. You will act as the central command during major incidents, coordinate cross-functional teams, drive incident triage and escalation, lead incident bridges and crisis calls, and ensure timely communications to stakeholders. You will track SLAs, conduct root cause analysis and post-incident reviews, and continuously improve incident management processes, playbooks, and runbooks while supporting audits and service improvement initiatives.
Required Skills
Business, Sales & Management
Project ManagementProgress TrackingRisk Management
Engineering, Construction & Trades
Geotechnical Engineering
Information Technology
DevOpsCloud ArchitectureIncident ManagementServiceNowJiraModel MonitoringLoggingITILAzureEnterprise ApplicationsTicketing SystemsTechnical Documentation
Soft Skills & Professional Competencies
Root Cause AnalysisStakeholder Management
Requirements
5+ years of experience in IT operations, cloud, or infrastructure roles.2+ years of experience in Incident or Major Incident Management.ITIL Foundation or ITIL Intermediate (Incident Management) certification preferred.Cloud certifications (AWS, Azure, OCI) are a plus.Strong understanding of cloud platforms (AWS, Azure, OCI) and Private cloud operations.Familiarity with monitoring, alerting, and logging tools.Good understanding of infrastructure components (compute, storage, networking, IAM).Ability to assess technical impact and prioritize incidents effectively.Experience with ITSM tools (ServiceNow, Jira, Remedy, etc.).Strong knowledge of ITIL Incident and Major Incident Management processes.
Description
The Incident Manager is responsible for overseeing the end-to-end management of incidents impacting cloud and infrastructure services, including AWS, Azure, and OCI environments. This role ensures rapid restoration of services, effective communication with stakeholders, and continuous improvement through post-incident analysis.Responsibilities:Own and manage the full incident lifecycle from detection to closure.Act as the central command point during major (P1/P2) incidents.Coordinate cross-functional teams including cloud, network and Infrastructure teams as well as CSMs.Ensure timely incident triage, escalation, and resolution.Lead incident bridges, war rooms, and crisis calls.Ensure accurate and timely communication to stakeholders and leadership.Track incidents against SLAs and ensure compliance with operational targets.Drive root cause analysis (RCA) and post-incident reviews (PIRs).Identify recurring issues and recommend preventive and corrective actions.Maintain and improve incident management processes, playbooks, and runbooks.Ensure proper documentation and ticket updates in ITSM tools.Support audits, reporting, and service improvement initiatives.
✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00