Peraton Logo

Peraton

Technical Enterprise Incident Manager

Posted 2 Days Ago
Be an Early Applicant
Remote
Hiring Remotely in United States
86K-138K Annually
Senior level
Remote
Hiring Remotely in United States
86K-138K Annually
Senior level
Leads enterprise major incident response, coordinates cross-functional teams during outages, drives service restoration, and provides executive communications. Improves cloud platform reliability through monitoring, observability, resiliency, automation, and root-cause remediation. Maintains incident procedures, runbooks, ServiceNow documentation, and operational metrics while supporting ITIL and SRE practices. Requires participation in an after-hours and weekend on-call rotation.
The summary above was generated by AI
Responsibilities

We are seeking a highly motivated and technically skilled Technical Enterprise Incident Manager with strong Cloud Platform DevSecOps Engineering and application experience to lead enterprise incident response, service restoration efforts, and operational reliability initiatives. This individual will serve as the central point of coordination during major incidents, ensuring rapid resolution, clear communication, and continuous service improvement across enterprise infrastructure and applications. 

The ideal candidate possesses a strong operational background, excellent communication skills, and hands-on technical expertise in infrastructure, cloud technologies, monitoring, automation, and IT service management processes. This role requires the ability to drive incident response while also identifying systemic reliability improvements.This position may require participation in an after-hours and weekend on-call rotation supporting enterprise production incidents and critical outage management activities. 


Key Responsibilities:


Enterprise Incident Management

  • Lead and coordinate Incident bridge calls involving infrastructure, application, network, cloud, security, and vendor teams.
  • Drive rapid service restoration while maintaining accurate timelines, communications, and executive updates.
  • Ensure incidents are prioritized appropriately based on business impact and operational risk.
  • Manage escalation procedures and engage leadership when required.
  • Monitor SLA compliance and ensure incident response metrics are consistently achieved.

Cloud Platform DevSecOps Engineering

  • Improve platform reliability, availability, observability, and operational maturity.
  • Work with application teams to facilitate issues and implement root cause remediations.
  • Develop and enhance monitoring, alerting, and dashboarding capabilities.
  • Analyze trends, KPIs, and operational metrics to proactively identify reliability risks.
  • Support implementation of resiliency strategies including redundancy, failover, capacity planning, and performance optimization.
  • Create and maintain cloud architecture and service dependency diagrams using Cloudcraft.
  • Utilize Datadog for monitoring, alert correlation, dashboards, incident investigation, and performance analysis.
  • Assist with production readiness reviews and operational acceptance activities.
  • Participate in after-hours on-call incident management rotation as required.

Operational Excellence

  • Develop and maintain incident management procedures, runbooks, and knowledge articles.
  • Ensure accurate ticket documentation within ServiceNow.
  • Drive continual service improvement initiatives aligned with ITIL and SRE best practices.
  • Collaborate with cross functional teams to improve communication, escalation paths, and operational workflows.
  • Support audit, compliance, and operational reporting requirements. 
Qualifications

Required Qualifications:

  • Bachelor’s degree and 5 years of experience or 9 years with a Highschool diploma.
  • At least 5 years of experience in Cloud Incident Management, Operations Engineering, NOC, SRE, Application or Production Support environments.
  • Experience leading enterprise Major Incident response efforts in a 24x7 operational environment.
  • Strong understanding of ITIL Incident and Problem Management processes.
  • Hands-on experience with infrastructure technologies including:
    • Windows/Linux Servers
    • Networking concepts
    • Cloud platforms (AWS, Azure, or GCP)
    • Load balancers, proxies, DNS, and firewalls
  • Experience with monitoring and observability platforms such as:
    • Datadog
    • Cloudcraft
    • CloudWatch
  • Experience using Cloudcraft to document and visualize cloud environments and application dependencies.
  • Experience using ServiceNow or similar ITSM platforms.
  • Strong analytical, troubleshooting, and organizational skills.
  • Excellent written and verbal communication skills with ability to facility meetings as well as brief technical teams and executive leadership.
  • Must be a US Citizen.
  • Must be able to obtain and maintain the required agency clearance.  

Preferred Qualifications:

  • Experience in a Site Reliability Engineering (SRE) or Cloud Platform DevOps environment.
  • Familiarity with CI/CD pipelines and Infrastructure as Code (IaC).
  • Experience supporting federal, healthcare, financial, or other highly regulated environments.
  • ITIL Foundation certification preferred.
  • SRE, cloud, or operational certifications are a plus.

 

 


Peraton Overview

Peraton is a next-generation national security company that drives missions of consequence spanning the globe and extending to the farthest reaches of the galaxy. As the world’s leading mission capability integrator and transformative enterprise IT provider, we deliver trusted, highly differentiated solutions and technologies to protect our nation and allies. Peraton operates at the critical nexus between traditional and nontraditional threats across all domains: land, sea, space, air, and cyberspace. The company serves as a valued partner to essential government agencies and supports every branch of the U.S. armed forces. Each day, our employees do the can’t be done by solving the most daunting challenges facing our customers. Visit peraton.com to learn how we’re keeping people around the world safe and secure.

Target Salary Range$86,000 - $138,000. This represents the typical salary range for this position. Salary is determined by various factors, including but not limited to, the scope and responsibilities of the position, the individual’s experience, education, knowledge, skills, and competencies, as well as geographic location and business and contract considerations. Depending on the position, employees may be eligible for overtime, shift differential, and a discretionary bonus in addition to base pay. EEOEEO: Equal opportunity employer, including disability and protected veterans, or other characteristics protected by law.

Similar Jobs

58 Minutes Ago
In-Office or Remote
152K-213K Annually
Expert/Leader
152K-213K Annually
Expert/Leader
Artificial Intelligence • Fintech • Information Technology • Logistics • Payments • Business Intelligence • Generative AI
Lead FP&A for global Value Services supporting Customer Success, Professional Services, Support, and related teams. Partner with the CCO, manage a team of three, own budgeting, forecasting, financial modeling, performance reporting, executive presentations, ad-hoc analysis, process improvements, headcount and spend controls to optimize margins and support strategic initiatives.
Top Skills: AnaplanChatgptCopilotGeminiExcelPowerPoint
2 Hours Ago
In-Office or Remote
Senior level
Senior level
Artificial Intelligence • Fintech • Software • Financial Services
Own enterprise new-logo acquisition and expansion while personally closing complex, high-value deals and carrying an individual quota. Hire, coach, and develop Enterprise AEs and SDRs; establish forecasting, pipeline, conversion, and sales operating cadences. Partner with Marketing, Product, Customer Success, Partnerships, Legal, and Finance, while presenting performance and strategic recommendations to executives and the board.
Top Skills: ChallengerChorusCommand Of The MessageForce ManagementGongHubspotLinkedin Sales NavigatorMeddpiccOutreachSalesloftZoominfo
2 Hours Ago
In-Office or Remote
Mid level
Mid level
Artificial Intelligence • Fintech • Software • Financial Services
Own vulnerability management across endpoints, servers, and cloud infrastructure; prioritize remediation and track SLA performance. Harden AWS environments, improve identity management, investigate cloud alerts, and support application security through SAST, DAST, dependency scanning, secure code reviews, and threat modeling. Resolve security tickets, document incidents and remediation, and lead security initiatives while collaborating with engineering, IT, and compliance.
Top Skills: AWSAws ConfigBashCloudtrailDastGuarddutyIamJIRALinearOwasp Top 10PythonQualysS3SastScaSecurity HubSnykSoc 2TenableVpcWiz

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account