1,433,937open jobs
84,184companies
217,995added this week
Browse all
Salary
≈ $39k – $99k per year (Estimated)
Location
Remote (Mexico)
Seniority
Staff · 8+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 10, 2026. First seen by Alion on Oct 8, 2026. Cognizant scores B on the Alion truth index.

Overview
Company
Impact
Profile match
Cognizant is an American information technology services and consulting company founded in 1994 as an in-house technology unit of Dun and Bradstreet in Chennai, India, and headquartered in Teaneck, New Jersey. It builds, integrates and operates software and infrastructure for large enterprises, with particularly deep positions in healthcare, financial services, insurance and life sciences, delivered by a global workforce of more than three hundred thousand people concentrated in India. Listed on Nasdaq, the company competes with the large Indian and global systems integrators and has reoriented its offerings around cloud migration, data platforms and AI-assisted engineering.

We’re hiring!

At Cognizant we have an ideal opportunity for you to be part of one of the largest companies in the digital sector worldwide. A Great Place To Work where we look for people who contribute new ideas, experiencing a dynamic and growing environment. At Cognizant we promote an inclusive culture, where we value different perspectives providing career growth and development opportunities. #WelcomeToCognizant!

Role Overview

The Application Reliability Engineer will lead a structured reliability-improvement initiative across critical applications. The consultant will analyze recurring production issues, identify systemic weaknesses, and drive measurable improvements in stability, resiliency, performance, observability, and operational effectiveness. This role requires a strong application-centric mindset, deep diagnostic skills, and the ability to convert operational noise into actionable reliability outcomes

Key Responsibilities

Reliability & Production Stability

· Lead reliability uplift initiatives across assigned applications and services.

· Assess current reliability posture: recurring incidents, degradation patterns, operational pain points, manual toil, and risk areas.

· Define and prioritize improvements that reduce incidents, customer impact, alert noise, and MTTR. Build a practical roadmap for resilience, availability, recoverability, and operational readiness. Evaluate operational procedures and identify gaps affecting responsiveness and production support. Partner with engineering, operations, platform, infrastructure, and product teams to execute the reliability plan.

Incident Management, Analysis & Root Cause Resolution

· Lead production incident reviews for high-impact and complex issues.

· Perform deep-dive analysis to identify root causes, not just symptoms.

· Review logs, metrics, traces, dependency health, and post-incident findings to identify systemic failure patterns.

· Differentiate actionable incidents from duplicates or noisy alerts.

· Facilitate RCA, corrective actions, and preventive-action planning.

· Establish traceability for incidents, remediation items, owners, due dates, and decisions.

· Track remediation through completion and validate effectiveness.

· Improve incident-management workflows: triage, escalation, response, resolution, and communication.

Operational Automation & Process Improvement

· Identify repetitive manual tasks causing delays, inconsistencies, or operational risk.

· Recommend and drive automation for support activities, health checks, data collection, remediation steps, and reporting.

· Improve workflows to reduce manual effort and increase consistency.

· Define process improvements to enhance reliability, predictability, change readiness, and service support.

Establish clear ownership and workflow visibility across teams.

· Application Resiliency & Design Improvement

· Evaluate applications for resiliency across dependencies, integrations, data flows, error handling, retry logic, timeouts, and recovery mechanisms.

· Recommend improvements to strengthen fault tolerance and reduce single points of failure.

· Assess behavior under degraded conditions and propose graceful-failure enhancements.

· Identify risks affecting reliability, availability, performance, or customer experience.

· Work with technical leads to embed reliability principles into enhancements and remediation work.

Performance & Application Tuning

· Identify performance bottlenecks across application layers, integrations, and infrastructure. Recommend tuning strategies for throughput, latency, resource utilization, and scalability.

· Validate improvements through load testing, profiling, and performance diagnostics.

· Collaborate with engineering teams to ensure performance readiness for peak loads.

· Monitoring, Alerting & Observability

· Assess current monitoring and alerting effectiveness.

· Define meaningful health indicators, SLIs, SLOs, and error budgets.

· Reduce alert noise and improve signal-to-noise ratio.

· Strengthen end-to-end observability across logs, metrics, traces, dashboards, and synthetic monitoring.

· Ensure monitoring coverage for critical workflows, dependencies, and failure modes.

· Stakeholder Leadership & Communication

· Act as the reliability lead for cross-functional teams.

· Communicate reliability posture, risks, priorities, and progress to leadership.

· Coordinate remediation activities across engineering, operations, and product teams.

· Provide structured updates, dashboards, and reliability metrics to stakeholders.

· Drive accountability for reliability commitments and timelines.

Required Skills & Qualifications

· 8+ years in application reliability, production engineering, SRE, or application operations.

· Strong diagnostic skills across logs, metrics, traces, and distributed systems.

· Experience with incident management, RCA, and operational readiness.

· Hands-on experience with monitoring tools (e.g., Splunk, Datadog, New Relic, AppDynamics, Grafana). Strong understanding of application architecture, integrations, APIs, databases, and cloud platforms. Experience with automation (Python, Shell, Ansible, Jenkins, or similar).

· Excellent communication and stakeholder-management skills.

· Ability to lead cross-team reliability initiatives independently.

Why Cognizant?

Improve your career in one of the largest and fastest growing IT services providers worldwide

Receive ongoing support and funding with training and development plans

Have a highly competitive benefits and salary package

Get the opportunity to work for leading global companies

We are committed to respecting human rights and build a better future by helping your minds and the environment

We invest in people and their wellbeing.

We create conditions for everyone to thrive. We do not discriminate based on race, religion, color, sex, age, disability, nationality, sexual orientation, gender identity or expression, or for any other reason covered.

Igualdad de Empleo y Política de Acción Afirmativa:

Cognizant es un empleador que ofrece igualdad de oportunidades. Todos los solicitantes calificados recibirán consideración para el empleo sin distinción de sexo, identidad de género, orientación sexual, raza, color, religión, origen nacional, discapacidad, estado de veterano protegido, edad o cualquier otra característica protegida por la ley.

At Cognizant we believe than our culture make us stronger!

Join us now!

#BeCognizant #IntuitionEngineered

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,433,937 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Mexico
AWS Cloud Engineer 23 min ago
$104k – $166k per year • Remote (United States) • Public Trust • 8+ years exp • High School Diploma
Python
JavaScript
PowerShell
Node JS
DevOps
Terraform
Ansible
CloudFormation
GitLab CI
CI/CD
Git
AWS
Docker
Kubernetes
Grafana
Gitflow
Configuration Management
Incident Management
IAM
Amazon ECS
Amazon CloudWatch
Management
Agile
Apply
Sr. Cloud Engineer 1 day ago
$104k – $138k per year • Remote (United States) • Full-Time • 7+ years exp • Bachelor's Degree • United States
Python
PowerShell
DevOps
Terraform
GitHub Actions
Datadog
CI/CD
GitOps
Jenkins
AWS
Docker
Kubernetes
Platform Engineering
Amazon EKS
IAM
API Gateway
VPN
Apply
≈ $109k – $201k per year (Estimated) • Remote (United States) • Full-Time • 5+ years exp • Bachelor's Degree • Chapel Hill
Management
OneDrive
SharePoint
Agile
ITIL
Apply
IT Systems Engineer 19 days ago
Remote (Greece) • Full-Time • 3+ years exp
PowerShell
DevOps
Azure
Windows Server
Incident Management
Windows
TCP/IP
DNS
DHCP
VPN
Cybersecurity
ISO 27001
Microsoft Defender
SOC 2
Least Privilege
Microsoft Entra ID
Management
Confluence
Jira
Apply
$100k – $146k per year • Equity • Remote (United States) • Full-Time • 3+ years exp • United States
PowerShell
DevOps
Windows Server
Incident Management
DNS
Cybersecurity
Active Directory
Apply
≈ $16k – $43k per year (Estimated) • In office • Full-Time • Novosibirsk
Python
Rust
Python
Django
Ruff
Databases
PostgreSQL
Redis
DevOps
Rest API
Ansible
Git
Cybersecurity
Keycloak
LDAP
QA
Pytest
Apply
≈ $12k – $29k per year (Estimated) • In office • Novosibirsk
Python
Java
C++
Bash
Databases
PostgreSQL
Cassandra
Apache Kafka
Apache Ignite
AI/ML
AI Agents
GigaChat
DevOps
Ansible
Zabbix
OpenShift
HAProxy
CI/CD
Jenkins
Kubernetes
Nginx
Grafana
Bitbucket
Linux
Apache HTTP Server
Management
Telegram
Apply
In office • Internship • Bachelor's Degree • Nanjing
Python
DevOps
Linux
TCP/IP
DHCP
Wi-Fi
Apply
≈ $37k – $102k per year (Estimated) • In office • Contractor • Associate's Degree • Singapore
Python
JavaScript
TypeScript
SQL
Node JS
Python
Flask
Django
AI/ML
Pandas
NumPy
RAG
Frontend
Vue.js
Angular
React.js
DevOps
Terraform
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Apply
In office • Full-Time • Bachelor's Degree • Singapore
Python
Java
DevOps
CI/CD
Analytics
ETL/ELT
AWS Glue
Management
Agile
Scrum
Apply
Remote (Brazil) • Full-Time • Bachelor's Degree • Brazil
Python
Bash
DevOps
Splunk
Datadog
Dynatrace
Prometheus
Azure
Grafana
Incident Management
Linux
Cybersecurity
PCI DSS
Apply
$113k – $130k per year • Equity • Remote (United States) • Full-Time • 5+ years exp • Mesa
C#
C#
ASP.NET Core
Databases
MS SQL
DevOps
Terraform
Helm
Azure DevOps
Azure
GitOps
Kubernetes
Bicep
Azure AKS
Cybersecurity
Microsoft Entra ID
QA
Rest-Assured
Apply
≈ $129k – $251k per year (Estimated) • Equity • Remote (United States) • Full-Time • 8+ years exp • Des Moines
Python
PowerShell
DevOps
Splunk
Terraform
Ansible
GitHub Actions
VMWare
Datadog
GitLab CI
CI/CD
Windows Server
Jenkins
AWS
Configuration Management
Amazon EC2
Amazon S3
IAM
Amazon CloudWatch
Windows
DNS
VPN
Cybersecurity
HIPAA
FedRAMP
Zero Trust
Active Directory
Apply
Remote (India) • Full-Time • Gurgaon
Python
PowerShell
DevOps
Splunk
Terraform
GitHub Actions
Prometheus
Azure
CI/CD
Kubernetes
Grafana
Azure AKS
Incident Management
DNS
Apply
$105k – $115k per year • Equity • Remote (United States) • Full-Time • Des Moines
DevOps
Self-Healing
AIOps
Management
ServiceNow
ITIL
ITSM
Apply
Field Engineer 1 day ago
≈ $30k – $83k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Mexico City
DevOps
Linux
Apply
$17k per year • In office • Contractor • Mexico
Apply
Analista de Costos 1 day ago
≈ $22k – $63k per year (Estimated) • In office • 4+ years exp • Bachelor's Degree • Mexico
Apply
CONTROLS ENGINEER 1 day ago
≈ $30k – $66k per year (Estimated) • In office • 4+ years exp • Bachelor's Degree • Mexico
Apply
≈ $66k – $153k per year (Estimated) • In office • Full-Time • 5+ years exp • Mexico
Python
Ruby
PowerShell
Bash
DevOps
Splunk
GCP
Azure
AWS
Linux
Windows
TCP/IP
Cybersecurity
Crowdstrike
Microsoft Sentinel
Microsoft Defender
GDPR
IBM QRadar
SIEM
Apply
See all jobs
This is one of many
1,433,937 more open roles from verified company boards, updated every day.