{"id":1471768,"url":"https://alion.io/job/peraton-technical-enterprise-incident-manager-2","title":"Technical Enterprise Incident Manager","company":{"id":722,"name":"Peraton","domain":"peraton.com","url":"https://alion.io/company/peraton","size_band":"1001-5000","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"iCIMS","truth_index":null},"role":"Security","role_family":"Security","seniority":"senior","employment_type":null,"work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["United States"],"countries":["US"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":86000,"max":138000,"currency":"USD","period":"year","gross":null,"usd_annual":138000},"salary_estimate":null,"experience_years_min":5,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AWS","optional":false},{"name":"Datadog","optional":false},{"name":"Incident Management","optional":false},{"name":"ITIL","optional":false},{"name":"ITSM","optional":false},{"name":"ServiceNow","optional":false},{"name":"SLI/SLO/SLA","optional":false},{"name":"Azure","optional":true},{"name":"DNS","optional":true},{"name":"GCP","optional":true},{"name":"Linux","optional":true},{"name":"Windows","optional":true}],"status":"live","first_seen_at":"2026-09-29T16:27:44Z","employer_posted_date":"2026-09-29","last_verified_at":"2026-09-30T11:39:03Z","board_verified":true,"closed_at":null,"days_open":2,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":2},"description":"Responsibilities\nPeraton is seeking a highly motivated and technically skilled Technical Enterprise Incident Manager with strong Cloud Platform and application experience to lead enterprise incident response, service restoration efforts, and operational reliability initiatives. This individual will serve as the central point of coordination during major incidents, ensuring rapid resolution, clear communication, and continuous service improvement across enterprise infrastructure and applications.\nThe ideal candidate possesses a strong operational background, excellent communication skills, and hands-on technical expertise in infrastructure, cloud technologies, monitoring, automation, and IT service management processes. This role requires the ability to drive incident response while also identifying systemic reliability improvements.\nLocation: Remote Shift Schedule: 8am - 5pm Eastern Standard Time (EST). This position will also participate in 24x7 on-call rotations for incident management.\nWhat You Will Do\nEnterprise Incident Management\nLead and coordinate Incident bridge calls involving infrastructure, application, network, cloud, security, and vendor teams.\nDrive rapid service restoration while maintaining accurate timelines, communications, and executive updates.\nEnsure incidents are prioritized appropriately based on business impact and operational risk.\nManage escalation procedures and engage leadership when required.\nMonitor SLA compliance and ensure incident response metrics are consistently achieved.\nFacilitate Post-Incident Reviews (PIRs) and ensure high-quality Root Cause Analysis (RCA) documentation and follow-through on corrective actions.\nDrive automation of incident detection, triage, and response workflows to reduce operational toil and improve resolution times.\nMaintain structured communication frameworks during major incidents, including timely stakeholder notifications, ongoing updates, and final incident summaries.\nValidate runbooks, service dependency maps, and technical documentation to ensure accuracy and usability during incidents.\nUtilize additional observability tools such as log aggregation and APM to enhance troubleshooting and incident analysis.\nCloud Platform DevSecOps Engineering\nWork with application teams to facilitate issues and implement root cause remediations.\nDevelop and enhance monitoring, alerting, and dashboarding capabilities.\nAnalyze trends, KPIs, and operational metrics to proactively identify reliability risks.\nSupport implementation of resiliency strategies including redundancy, failover, capacity planning, and performance optimization.\nUtilize Datadog for monitoring, alert correlation, dashboards, incident investigation, and performance analysis.\nParticipate in after-hours on-call incident management rotation as required.\nOperational Excellence\nDevelop and maintain incident management procedures, runbooks, and knowledge articles.\nEnsure accurate ticket documentation within ServiceNow.\nDrive continual service improvement initiatives aligned with ITIL and SRE best practices.\nCollaborate with cross functional teams to improve communication, escalation paths, and operational workflows.\nSupport audit, compliance, and operational reporting requirements. \nQualifications\nBasic Qualifications:\nMust be a U.S. citizen with the ability to obtain and maintain the required Public Trust level clearance\nBachelor’s Degree and 5 years of experience, or a High School diploma or equivalent and 9 years of experience\n5+ years of experience in Cloud Incident Management, Operations Engineering, NOC, SRE, Application or Production Support environments.\nExperience leading enterprise Major Incident response efforts in a 24x7 operational environment.\nStrong understanding of ITIL Incident and Problem Management processes.\n3+ years of experience working with AWS cloud services\nExperience with monitoring and observability platforms such as Datadog, Cloudcraft, or similar\nExperience using ServiceNow or similar ITSM platforms.\nStrong analytical, troubleshooting, and organizational skills.\nExcellent written and verbal communication skills with ability to facility meetings as well as brief technical teams and executive leadership.\nPreferred Qualifications:\nExperience in a Site Reliability Engineering (SRE) or Cloud Platform DevOps environment.\nExperience supporting federal, healthcare, financial, or other highly regulated environments.\nHands-on experience with infrastructure technologies including:\nWindows/Linux Servers\nNetworking concepts\nCloud platforms (AWS, Azure, or GCP)\nLoad balancers, proxies, DNS, and firewalls\nPeraton Overview\nPeraton is a next-generation national security company that drives missions of consequence spanning the globe and extending to the farthest reaches of the galaxy. As the world’s leading mission capability integrator and transformative enterprise IT provider, we deliver trusted, highly differentiated solutions and technologies to protect our nation and allies. Peraton operates at the critical nexus between traditional and nontraditional threats across all domains: land, sea, space, air, and cyberspace. The company serves as a valued partner to essential government agencies and supports every branch of the U.S. armed forces. Each day, our employees do the can’t be done by solving the most daunting challenges facing our customers. Visit peraton.com to learn how we’re keeping people around the world safe and secure.\nTarget Salary Range\n$86,000 - $138,000. This represents the typical salary range for this position. Salary is determined by various factors, including but not limited to, the scope and responsibilities of the position, the individual’s experience, education, knowledge, skills, and competencies, as well as geographic location and business and contract considerations. Depending on the position, employees may be eligible for overtime, shift differential, and a discretionary bonus in addition to base pay.EEO\nEEO: Equal opportunity employer, including disability and protected veterans, or other characteristics protected by law.","description_format":"text","description_chars":6086,"description_truncated":false,"requirements":{"experience_years_min":5,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"high_school","optional":true},"security_clearance":true,"languages":[]},"benefits":[],"hiring_locations":[{"name":"United States","iso":"US","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["National Security","Homeland Security"],"lifecycle":[{"event":"open","at":"2026-09-29T16:27:44Z"}],"liveness":{"score":76,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.844,"p_room":0.9,"age_days":1,"expected_fill_days":3,"reasons":["conf:18","velocity","win:mid","comp:brand"],"computed_at":"2026-10-01T05:45:00Z"},"pay":{"stated_usd_annual":138000,"is_top_pay":false},"html_url":"https://alion.io/job/peraton-technical-enterprise-incident-manager-2","json_url":"https://alion.io/job/peraton-technical-enterprise-incident-manager-2.json","meta":{"generated_at":"2026-10-01T20:35:19Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":2986,"day_limit":5000,"remaining_today":2014,"minute_limit":60,"resets_at":"2026-10-02T00:00:00Z"}}}