Technical Enterprise Incident Manager

Peraton

Reston (VA)

On-site

USD 86,000 - 138,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Peraton is looking for a Technical Enterprise Incident Manager in Reston, Virginia, to lead incident response and support operational reliability. The ideal candidate will have a strong background in Cloud technologies and ITIL practices, along with excellent communication and analytical skills.

The position requires overseeing major incidents, ensuring timely restoration, and developing operational protocols. A Bachelor’s degree or 9 years of experience is required. The salary range is $86,000 - $138,000.

Qualifications

  • 5 years of experience in Cloud Incident Management.
  • Leading major incident response in a 24x7 environment.
  • Experience with infrastructure technologies including Windows/Linux Servers.

Responsibilities

  • Lead incident bridge calls for coordination during incidents.
  • Drive rapid service restoration and monitor SLA compliance.
  • Develop and maintain incident management procedures and runbooks.

Skills

Technical expertise in Cloud technologies
Incident Management
ITIL Incident and Problem Management
Strong communication skills
Analytical skills

Education

Bachelor’s degree
9 years experience with a high school diploma

Tools

ServiceNow
Datadog
Cloudcraft
AWS
Azure
GCP

Job description

Responsibilities

We are seeking a highly motivated and technically skilled Technical Enterprise Incident Manager with strong Cloud Platform DevSecOps Engineering and application experience to lead enterprise incident response, service restoration efforts, and operational reliability initiatives. This individual will serve as the central point of coordination during major incidents, ensuring rapid resolution, clear communication, and continuous service improvement across enterprise infrastructure and applications.

The ideal candidate possesses a strong operational background, excellent communication skills, and hands‑on technical expertise in infrastructure, cloud technologies, monitoring, automation, and IT service management processes. This role requires the ability to drive incident response while also identifying systemic reliability improvements.

This position may require participation in an after‑hours and weekend on‑call rotation supporting enterprise production incidents and critical outage management activities.

Key Responsibilities
  • Lead and coordinate Incident bridge calls involving infrastructure, application, network, cloud, security, and vendor teams.
  • Drive rapid service restoration while maintaining accurate timelines, communications, and executive updates.
  • Ensure incidents are prioritized appropriately based on business impact and operational risk.
  • Manage escalation procedures and engage leadership when required.
  • Monitor SLA compliance and ensure incident response metrics are consistently achieved.
Cloud Platform DevSecOps Engineering
  • Improve platform reliability, availability, observability, and operational maturity.
  • Work with application teams to facilitate issues and implement root cause remediations.
  • Develop and enhance monitoring, alerting, and dashboarding capabilities.
  • Analyze trends, KPIs, and operational metrics to proactively identify reliability risks.
  • Support implementation of resiliency strategies including redundancy, failover, capacity planning, and performance optimization.
  • Create and maintain cloud architecture and service dependency diagrams using Cloudcraft.
  • Utilize Datadog for monitoring, alert correlation, dashboards, incident investigation, and performance analysis.
  • Assist with production readiness reviews and operational acceptance activities.
  • Participate in after‑hours on‑call incident management rotation as required.
Operational Excellence
  • Develop and maintain incident management procedures, runbooks, and knowledge articles.
  • Ensure accurate ticket documentation within ServiceNow.
  • Drive continual service improvement initiatives aligned with ITIL and SRE best practices.
  • Collaborate with cross‑functional teams to improve communication, escalation paths, and operational workflows.
  • Support audit, compliance, and operational reporting requirements.
Qualifications

Required Qualifications:

  • Bachelor’s degree and 5 years of experience or 9 years with a high school diploma.
  • At least 5 years of experience in Cloud Incident Management, Operations Engineering, NOC, SRE, Application or Production Support environments.
  • Experience leading enterprise major incident response efforts in a 24x7 operational environment.
  • Strong understanding of ITIL Incident and Problem Management processes.
  • Hands‑on experience with infrastructure technologies including Windows/Linux Servers, networking concepts, cloud platforms (AWS, Azure, or GCP), load balancers, proxies, DNS, and firewalls.
  • Experience with monitoring and observability platforms such as Datadog, Cloudcraft, CloudWatch.
  • Experience using Cloudcraft to document and visualize cloud environments and application dependencies.
  • Experience using ServiceNow or similar ITSM platforms.
  • Strong analytical, troubleshooting, and organizational skills.
  • Excellent written and verbal communication skills with ability to facilitate meetings as well as brief technical teams and executive leadership.
  • Must be a US Citizen.
  • Must be able to obtain and maintain the required agency clearance.

Preferred Qualifications:

  • Experience in a Site Reliability Engineering (SRE) or Cloud Platform DevOps environment.
  • Familiarity with CI/CD pipelines and Infrastructure as Code (IaC).
  • Experience supporting federal, healthcare, financial, or other highly regulated environments.
  • ITIL Foundation certification preferred.
  • SRE, cloud, or operational certifications are a plus.
Target Salary Range

$86,000 - $138,000. This represents the typical salary range for this position. Salary is determined by various factors, including but not limited to, the scope and responsibilities of the position, the individual's experience, education, knowledge, skills, and competencies, as well as geographic location and business and contract considerations. Depending on the position, employees may be eligible for overtime, shift differential, and a discretionary bonus in addition to base pay.

EEO

Equal opportunity employer, including disability and protected veterans, or other characteristics protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Technical Enterprise Incident Manager
Technical Enterprise Incident Manager

Peraton • United States

On-site
USD 86,000 - 138,000
Medical and dental insurance
401(k) plan
Paid time off (PTO)
Incident Manager
Incident Manager

Aceolution • Pennsylvania

On-site
USD 120,000 - 150,000
Enterprise Operations Senior Analyst
Enterprise Operations Senior Analyst

Sidley Austin • Chicago (IL)

Hybrid
USD 102,000 - 118,000
Enterprise Operations Senior Analyst
Enterprise Operations Senior Analyst

LLP • Chicago (IL)

Hybrid
USD 102,000 - 118,000
Incident Management Analyst
Incident Management Analyst

Insight Global • New York (NY)

On-site
USD 90,000 - 130,000
Sr. Associate, Incident Manager, Incident Management Hub (IMH)
Sr. Associate, Incident Manager, Incident Management Hub (IMH)

Santander US • Quincy (MA)

On-site
USD 120,000 - 160,000
Major Incident Manager
Major Incident Manager

Belcan LLC • Iowa City (IA)

On-site
USD 110,000 - 125,000
Health care
Dental
Vision
+4
Senior AIOps and Incident Management / Site Reliability Engineering C2C jobs
Senior AIOps and Incident Management / Site Reliability Engineering C2C jobs

Tech Mirrors • Fort Mill (SC)

Hybrid
USD 140,000 - 190,000
Enterprise SRE – Incident Coordinator
Enterprise SRE – Incident Coordinator

Jobtailor • Missouri

On-site
USD 90,000 - 120,000
Manager-Cloud Operations
Manager-Cloud Operations

WellSpan Health • York

On-site
USD 90,000 - 120,000
Comprehensive health benefits
Retirement savings plan
Paid time off (PTO)
+3