Location
In office
Seniority
Senior
Confirmed on the employer's own hiring board on Sep 25, 2026. First seen by Alion on Sep 25, 2026.
Overview
Company
Impact
Profile match
HCLTech is a major Indian multinational information technology (IT) services and consulting company headquartered in Noida, Uttar Pradesh. Spun off from the original HCL Group in 1991, it ranks as one of India's largest technology companies alongside firms like TCS, Infosys, and Wipro.
Job Summary
Incident Detection & Front-Line Triage
- Pattern Recognition: Monitor incoming service desk queues to spot sudden spikes in identical user complaints, indicating a potential widespread system outage.
- Alert Validation: Review automated infrastructure and application alerts to filter out false positives before triggering the major incident workflow.
- Impact Assessment: Interview affected business units quickly to map out the scope, user count, and financial implications of an ongoing disruption.
- Ticket Categorization: Ensure all major incident tickets are accurately tagged with the correct high-priority urgency (P1/P2) and functional categories.
- Process Activation: Initiate the formal Major Incident Management (MIM) workflow immediately upon identifying breaches of critical service-level thresholds.
- War Room Participation: Join live technical crisis bridges to provide real-time updates and collaborate dynamically with cross-functional technical teams. Service Restoration & Remediation
- Workaround Deployment: Apply pre-approved temporary workarounds or failover mechanisms to restore business continuity ahead of a permanent fix.
- Validation Testing: Conduct comprehensive smoke tests and user acceptance testing to confirm the systems are fully operational before closing the incident. Communication, Collaboration & Post-Incident Actions
- Stakeholder Notification: Broadcast standardized, jargon-free status updates to impacted end-users and executive leadership at regular intervals.
- Vendor Coordination: Escalate tickets to third-party providers or external software vendors and track their progress against strict contractual SLAs.
- L3 Escalation: Package all technical diagnostic data, logs, and troubleshooting steps neatly when escalating unresolved issues to Level 3 engineering teams.
- Chronology Tracking: Maintain a meticulous, minute-by-minute timeline of technical actions taken, system behaviors, and milestones throughout the live incident lifecycle.
- PIR Contribution: Provide technical root-cause data and timeline logs to the Major Incident Manager for the formal Post-Incident Review (PIR) and Problem Management records. Incident Detection & Front-Line Triage
- Pattern Recognition: Monitor incoming service desk queues to spot sudden spikes in identical user complaints, indicating a potential widespread system outage.
- Alert Validation: Review automated infrastructure and application alerts to filter out false positives before triggering the major incident workflow.
- Impact Assessment: Interview affected business units quickly to map out the scope, user count, and financial implications of an ongoing disruption.
- Ticket Categorization: Ensure all major incident tickets are accurately tagged with the correct high-priority urgency (P1/P2) and functional categories.
- Process Activation: Initiate the formal Major Incident Management (MIM) workflow immediately upon identifying breaches of critical service-level thresholds.
- War Room Participation: Join live technical crisis bridges to provide real-time updates and collaborate dynamically with cross-functional technical teams. Service Restoration & Remediation
- Workaround Deployment: Apply pre-approved temporary workarounds or failover mechanisms to restore business continuity ahead of a permanent fix.
- Validation Testing: Conduct comprehensive smoke tests and user acceptance testing to confirm the systems are fully operational before closing the incident. Communication, Collaboration & Post-Incident Actions
- Stakeholder Notification: Broadcast standardized, jargon-free status updates to impacted end-users and executive leadership at regular intervals.
- Vendor Coordination: Escalate tickets to third-party providers or external software vendors and track their progress against strict contractual SLAs.
- L3 Escalation: Package all technical diagnostic data, logs, and troubleshooting steps neatly when escalating unresolved issues to Level 3 engineering teams.
- Ch
Key Responsibilities
NA
Job Description : Incident Detection & Front-Line Triage
- Pattern Recognition: Monitor incoming service desk queues to spot sudden spikes in identical user complaints, indicating a potential widespread system outage.
- Alert Validation: Review automated infrastructure and application alerts to filter out false positives before triggering the major incident workflow.
- Impact Assessment: Interview affected business units quickly to map out the scope, user count, and financial implications of an ongoing disruption.
- Ticket Categorization: Ensure all major incident tickets are accurately tagged with the correct high-priority urgency (P1/P2) and functional categories.
- Process Activation: Initiate the formal Major Incident Management (MIM) workflow immediately upon identifying breaches of critical service-level thresholds.
- War Room Participation: Join live technical crisis bridges to provide real-time updates and collaborate dynamically with cross-functional technical teams.
Service Restoration & Remediation
- Workaround Deployment: Apply pre-approved temporary workarounds or failover mechanisms to restore business continuity ahead of a permanent fix.
- Validation Testing: Conduct comprehensive smoke tests and user acceptance testing to confirm the systems are fully operational before closing the incident.
Communication, Collaboration & Post-Incident Actions
- Stakeholder Notification: Broadcast standardized, jargon-free status updates to impacted end-users and executive leadership at regular intervals.
- Vendor Coordination: Escalate tickets to third-party providers or external software vendors and track their progress against strict contractual SLAs.
- L3 Escalation: Package all technical diagnostic data, logs, and troubleshooting steps neatly when escalating unresolved issues to Level 3 engineering teams.
- Chronology Tracking: Maintain a meticulous, minute-by-minute timeline of technical actions taken, system behaviors, and milestones throughout the live incident lifecycle.
- PIR Contribution: Provide technical root-cause data and timeline logs to the Major Incident Manager for the formal Post-Incident Review (PIR) and Problem Management records.
Skill Requirements
Skill Requirement : Incident Detection & Front-Line Triage
- Pattern Recognition: Monitor incoming service desk queues to spot sudden spikes in identical user complaints, indicating a potential widespread system outage.
- Alert Validation: Review automated infrastructure and application alerts to filter out false positives before triggering the major incident workflow.
- Impact Assessment: Interview affected business units quickly to map out the scope, user count, and financial implications of an ongoing disruption.
- Ticket Categorization: Ensure all major incident tickets are accurately tagged with the correct high-priority urgency (P1/P2) and functional categories.
- Process Activation: Initiate the formal Major Incident Management (MIM) workflow immediately upon identifying breaches of critical service-level thresholds.
- War Room Participation: Join live technical crisis bridges to provide real-time updates and collaborate dynamically with cross-functional technical teams. Service Restoration & Remediation
- Workaround Deployment: Apply pre-approved temporary workarounds or failover mechanisms to restore business continuity ahead of a permanent fix.
- Validation Testing: Conduct comprehensive smoke tests and user acceptance testing to confirm the systems are fully operational before closing the incident. Communication, Collaboration & Post-Incident Actions
- Stakeholder Notification: Broadcast standardized, jargon-free status updates to impacted end-users and executive leadership at regular intervals.
- Vendor Coordination: Escalate tickets to third-party providers or external software vendors and track their progress against strict contractual SLAs.
- L3 Escalation: Package all technical diagnostic data, logs, and troubleshooting steps neatly when escalating unresolved issues to Level 3 engineering teams.
- Chronology Tracking: Maintain a meticulous, minute-by-minute timeline of technical actions taken, system behaviors, and milestones throughout the live incident lifecycle.
- PIR Contribution: Provide technical root-cause data and timeline logs to the Major Incident Manager for the formal Post-Incident Review (PIR) and Problem Management records. Other Requirement : Incident Detection & Front-Line Triage
- Pattern Recognition: Monitor incoming service desk queues to spot sudden spikes in identical user complaints, indicating a potential widespread system outage.
- Alert Validation: Review automated infrastructure and application alerts to filter out false positives before triggering the major incident workflow.
- Impact Assessment: Interview affected business units quickly to map out the scope, user count, and financial implications of an ongoing disruption.
- Ticket Categorization: Ensure all major incident tickets are accurately tagged with the correct high-priority urgency (P1/P2) and functional categories.
- Process Activation: Initiate the formal Major Incident Management (MIM) workflow immediately upon identifying breaches of critical service-level thresholds.
- War Room Participation: Join live technical crisis bridges to provide real-time updates and collaborate dynamically with cross-functional technical teams. Service Restoration & Remediation
- Workaround Deployment: Apply pre-approved temporary workarounds or failover mechanisms to restore business continuity ahead of a permanent fix.
- Validation Testing: Conduct comprehensive smoke tests and user acceptance testing to confirm the systems are fully operational before closing the incident. Communication, Collaboration & Post-Incident Actions
- Stakeholder Notification: Broadcast standardized, jargon-free status updates to impacted end-users and executive leadership at regular intervals.
- Vendor Coordination: Escalate tickets to third-party providers or external software vendors and track their progress against strict contractual SLAs.
- L3 Escalation: Package all technical diagnostic data, logs, and troubleshooting steps neatly when escalating unresolved is
Other Requirements
Other Requirement : Incident Detection & Front-Line Triage
- Pattern Recognition: Monitor incoming service desk queues to spot sudden spikes in identical user complaints, indicating a potential widespread system outage.
- Alert Validation: Review automated infrastructure and application alerts to filter out false positives before triggering the major incident workflow.
- Impact Assessment: Interview affected business units quickly to map out the scope, user count, and financial implications of an ongoing disruption.
- Ticket Categorization: Ensure all major incident tickets are accurately tagged with the correct high-priority urgency (P1/P2) and functional categories.
- Process Activation: Initiate the formal Major Incident Management (MIM) workflow immediately upon identifying breaches of critical service-level thresholds.
- War Room Participation: Join live technical crisis bridges to provide real-time updates and collaborate dynamically with cross-functional technical teams. Service Restoration & Remediation
- Workaround Deployment: Apply pre-approved temporary workarounds or failover mechanisms to restore business continuity ahead of a permanent fix.
- Validation Testing: Conduct comprehensive smoke tests and user acceptance testing to confirm the systems are fully operational before closing the incident. Communication, Collaboration & Post-Incident Actions
- Stakeholder Notification: Broadcast standardized, jargon-free status updates to impacted end-users and executive leadership at regular intervals.
- Vendor Coordination: Escalate tickets to third-party providers or external software vendors and track their progress against strict contractual SLAs.
- L3 Escalation: Package all technical diagnostic data, logs, and troubleshooting steps neatly when escalating unresolved issues to Level 3 engineering teams.
- Chronology Tracking: Maintain a meticulous, minute-by-minute timeline of technical actions taken, system behaviors, and milestones throughout the live incident lifecycle.
- PIR Contribution: Provide technical root-cause data and timeline logs to the Major Incident Manager for the formal Post-Incident Review (PIR) and Problem Management records.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
751,436 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Free forever. No card. Under a minute.
Your match
How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.
Recommended for you based on this role
Industrial Engineering
Similar stack
Same company
In your city
Agricultural Supervisor
1 day ago
≈ $40k – $85k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Riyadh
Apply
Apply
Apply
Apply
Engineer - Renewable Energy & Energy Storage
2 hours ago
$70k – $75k per year • In office • Bachelor's Degree • Cortland
Management
Microsoft Office
Apply
IT Systems & Infrastructure Engineer
4 hours ago
≈ $17k – $41k per year (Estimated) • In office • 5+ years exp • Muntinlupa
DevOps
Azure
Incident Management
Windows
TCP/IP
DNS
DHCP
VPN
Cybersecurity
Active Directory
Management
Jira
ServiceNow
ITIL
Service Desk
Apply
IT Infrastructure Engineer
4 hours ago
≈ $17k – $41k per year (Estimated) • In office • 5+ years exp • Muntinlupa
DevOps
Azure
Incident Management
Windows
TCP/IP
DNS
DHCP
VPN
Cybersecurity
Active Directory
Management
Jira
ServiceNow
ITIL
Service Desk
Apply
Solution Architect (DFS)
3 hours ago
In office • 8+ years exp • Bachelor's Degree
Python
Databases
Google BigQuery
BigQuery
AI/ML
LangGraph
LangChain
Model Context Protocol
Prompt Engineering
Function Calling
AI Agents
AgentOps
Gemini
RAG
Google ADK
A2A
Human-in-the-Loop
LLM Guardrails
Multi-Agent Systems
Tool Use
DevOps
Rest API
Terraform
GCP
OpenTelemetry
CI/CD
Git
Docker
Kubernetes
Platform Engineering
Google GKE
Google Cloud Run
AIOps
Incident Management
IAM
Management
ServiceNow
ITSM
Apply
Solution Architect (DFS)
3 hours ago
In office • 8+ years exp • Bachelor's Degree
Python
Databases
Google BigQuery
BigQuery
AI/ML
LangGraph
LangChain
Model Context Protocol
Prompt Engineering
Function Calling
AI Agents
AgentOps
Gemini
RAG
Google ADK
A2A
Human-in-the-Loop
LLM Guardrails
Multi-Agent Systems
Tool Use
DevOps
Rest API
Terraform
GCP
OpenTelemetry
CI/CD
Git
Docker
Kubernetes
Platform Engineering
Google GKE
Google Cloud Run
AIOps
Incident Management
IAM
Management
ServiceNow
ITSM
Apply
SeniorAdministrator - ServiceNow,JavaScript
3 hours ago
In office
Python
JavaScript
Node JS
AI/ML
NLP
OpenAI
DevOps
Rest API
Azure
CI/CD
AWS
SLI/SLO/SLA
Management
ServiceNow
Outlook
Microsoft Teams
ITSM
Service Desk
Apply
Apply
Apply
Solution Architect (DFS)
3 hours ago
In office • 8+ years exp • Bachelor's Degree
Python
Databases
Google BigQuery
BigQuery
AI/ML
LangGraph
LangChain
Model Context Protocol
Prompt Engineering
Function Calling
AI Agents
AgentOps
Gemini
RAG
Google ADK
A2A
Human-in-the-Loop
LLM Guardrails
Multi-Agent Systems
Tool Use
DevOps
Rest API
Terraform
GCP
OpenTelemetry
CI/CD
Git
Docker
Kubernetes
Platform Engineering
Google GKE
Google Cloud Run
AIOps
Incident Management
IAM
Management
ServiceNow
ITSM
Apply
In office • 10+ years exp
Management
Confluence
SharePoint
Apply
Azure DevOps Technical Consultant
3 hours ago
In office
Python
PowerShell
DevOps
Azure DevOps
Azure
CI/CD
Kubernetes
Apply
This is one of many
751,436 more open roles from verified company boards, updated every day.

