1,011,843open jobs
60,160companies
168,355added this week
Browse all
Salary
≈ $12k – $30k per year (Estimated)
Location
In office (Chennai)
Seniority
Middle · 3+ years exp

Confirmed on the employer's own hiring board on Oct 1, 2026. First seen by Alion on Sep 30, 2026. Pearson scores B on the Alion truth index.

Overview
Company
Impact
Profile match
Pearson is a London-based learning company that publishes educational courseware, runs assessments and qualifications, and provides digital learning, English-language and workforce skills services. Founded in 1844 and listed in London and New York, it employs more than 20,000 people, and its UK Assessment and Qualifications arm delivers nearly four million exam results a year, including GCSE, A-Level and BTEC. It hires for software and data engineering, product management, channel sales, partnership management, payroll, HR operations and tutoring roles across India, the UK, the US, Poland, Spain, Israel and the Philippines.

Function: Observability Engineering / SRE / Cloud Engineering

About the Role

We are looking for an Observability Engineer to join our Observability Engineering team and help build, standardize, and operate a modern enterprise observability platform across cloud and application environments.

The ideal candidate will have strong hands-on engineering experience with New Relic, Grafana, and modern observability technologies, combined with solid foundations in cloud engineering, automation, DevOps, and SRE practices.

This role goes beyond dashboard creation and monitoring operations. You will be responsible for engineering scalable observability solutions, defining telemetry standards, building reusable monitoring capabilities, improving application and infrastructure visibility, and enabling engineering teams to adopt observability as part of their software delivery lifecycle.

You will work closely with SRE, Cloud Engineering, Platform Engineering, Application Engineering, and Operations teams to build a consistent and scalable observability experience across the organization.

Key Responsibilities

1. Observability Engineering

  • Design, implement, and maintain enterprise-grade observability solutions across applications, infrastructure, cloud platforms, and services.
  • Build and maintain monitoring, alerting, dashboards, service health views, and operational telemetry.
  • Develop standardized observability patterns for metrics, logs, traces, events, and application performance monitoring.
  • Implement observability solutions using New Relic, Grafana, and other industry-standard tools.
  • Develop reusable dashboards, alerts, instrumentation patterns, and observability components.
  • Establish observability standards and best practices across engineering teams.
  • Continuously improve signal quality by reducing alert noise, false positives, and non-actionable alerts.

2. New Relic Engineering

  • Hands-on engineering experience with New Relic APM, Infrastructure Monitoring, Browser Monitoring, Synthetic Monitoring, Logs, Distributed Tracing, NRQL, Alerts, Workloads and Dashboards.
  • Design and implement New Relic monitoring and alerting strategies for enterprise applications.
  • Develop complex NRQL queries, alert conditions, dashboards, and operational views.
  • Configure and optimize New Relic agents and integrations.
  • Implement application and infrastructure instrumentation.
  • Develop reusable New Relic configurations and automation using APIs/IaC where appropriate.
  • Participate in New Relic platform governance, licensing optimization, and standardization.
  • Evaluate and implement emerging New Relic capabilities to improve engineering productivity and reliability.

3. Grafana & Visualization

  • Build and maintain operational dashboards using Grafana.
  • Integrate Grafana with multiple telemetry and data sources.
  • Design effective dashboards for application health, infrastructure, SRE, NOC, and executive operational visibility.
  • Develop visualization standards and reusable dashboard templates.
  • Understand the difference between visualization, monitoring, alerting, and observability, and apply each appropriately.

4. OpenTelemetry & Modern Observability

  • Experience with OpenTelemetry and modern telemetry architectures.
  • Implement and manage telemetry collection for metrics, logs, and traces.
  • Understand distributed tracing and service dependency mapping.
  • Work with telemetry pipelines, collectors, agents, exporters, and integrations.
  • Experience with technologies such as Prometheus, Loki, Elastic, Splunk, Datadog, Dynatrace, AppDynamics, or similar observability platforms is desirable.
  • Evaluate new observability technologies and recommend solutions based on scalability, cost, reliability, and engineering value.

5. Cloud Engineering

Strong cloud engineering fundamentals are expected, including experience with one or more major cloud platforms:

  • AWS
  • Microsoft Azure
  • Google Cloud Platform

Experience should include:

  • Compute, networking, storage, databases, containers, and cloud-native services.
  • Cloud monitoring and logging.
  • IAM and security fundamentals.
  • Infrastructure automation.
  • Cloud-native architecture and operational best practices.
  • Troubleshooting distributed cloud environments.

6. Automation & Infrastructure as Code

  • Automate repetitive observability and operational activities.
  • Develop scripts and tools using Python, Bash, Go, or similar languages.
  • Use Terraform / OpenTofu or equivalent Infrastructure as Code technologies.
  • Build reusable automation for dashboards, alerts, instrumentation, integrations, and configuration management.
  • Integrate observability capabilities into CI/CD pipelines.

7. DevOps & CI/CD

  • Experience with modern CI/CD practices and tools.
  • Hands-on experience with GitHub Actions, Jenkins, GitLab CI, Azure DevOps, or similar platforms.
  • Integrate observability and quality gates into deployment pipelines.
  • Implement deployment markers and release health monitoring.
  • Enable automated validation of application and infrastructure health following deployments.

8. SRE & Reliability Engineering

  • Apply SRE principles to improve system reliability and operational maturity.
  • Define and monitor SLIs, SLOs, and error budgets.
  • Participate in incident investigation and root-cause analysis.
  • Develop proactive monitoring and reliability solutions.
  • Identify reliability gaps and engineer solutions to eliminate recurring incidents.
  • Support capacity, performance, availability, and resilience engineering.

Required Skills & Experience

Must Have

  • 3+ years of experience in Observability, SRE, DevOps, Cloud Engineering, Platform Engineering, or a related engineering discipline.
  • Strong hands-on experience with New Relic.
  • Strong hands-on experience with Grafana.
  • Experience building production-grade dashboards, monitoring and alerting solutions.
  • Strong understanding of APM, infrastructure monitoring, logging, metrics, tracing, and distributed systems.
  • Experience with NRQL and New Relic alerting.
  • Experience with OpenTelemetry is highly desirable.
  • Strong cloud engineering experience in AWS, Azure, or GCP.
  • Strong scripting/programming experience in Python, Bash, Go, or similar.
  • Experience with Terraform/OpenTofu or another Infrastructure-as-Code technology.
  • Understanding of CI/CD and DevOps practices.
  • Understanding of SRE principles, SLIs, SLOs, incident management, and reliability engineering.
  • Strong troubleshooting and analytical skills.

Good to Have

  • New Relic certifications or equivalent hands-on expertise.
  • Grafana/Prometheus experience.
  • OpenTelemetry implementation experience.
  • Kubernetes and container observability.
  • Experience with Prometheus, Loki, Elastic, Splunk, Datadog, Dynatrace, AppDynamics or similar platforms.
  • Experience designing enterprise observability architectures.
  • Experience with observability platform migrations or consolidation.
  • Experience with observability cost optimization and licensing governance.
  • Experience developing observability-as-code.
  • Experience integrating observability with ServiceNow, Jira, PagerDuty, Opsgenie, or similar ITSM/incident platforms.
  • Experience with AI-assisted observability, AIOps, anomaly detection, or automated incident investigation.

What Success Looks Like

In this role, you will be successful when you can:

  • Engineer rather than simply operate monitoring.
  • Build scalable observability solutions that can be reused across hundreds of applications.
  • Turn raw telemetry into meaningful engineering and operational insights.
  • Reduce alert noise and improve signal quality.
  • Standardize New Relic and Grafana adoption across engineering teams.
  • Automate observability configuration and reduce manual operational work.
  • Improve application reliability through better telemetry and proactive detection.
  • Enable engineering teams to own more of their operational health.
  • Help establish Observability as a Platform/Capability, rather than simply another monitoring tool.

Behavioral & Leadership Skills

  • Strong ownership and accountability.
  • Ability to work across engineering, application, infrastructure, and operations teams.
  • Strong communication and stakeholder management skills.
  • Ability to explain complex technical concepts to both engineers and non-technical stakeholders.
  • Strong problem-solving and analytical mindset.
  • Comfortable working in a fast-paced, enterprise engineering environment.
  • Passion for automation, engineering excellence, reliability, and continuous improvement.
  • Ability to challenge existing approaches and introduce better engineering practices.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,011,843 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Chennai
≈ $17k – $37k per year (Estimated) • In office • Full-Time • 10+ years exp • Pune
Groovy
Databases
MySQL
PostgreSQL
MariaDB
Apache Kafka
DevOps
Terraform
Ansible
GCP
GitLab CI
CI/CD
ArgoCD
Jenkins
Kubernetes
Nginx
Grafana
Google GKE
Apply
≈ $18k – $41k per year (Estimated) • In office • 6+ years exp • Bachelor's Degree • Bengaluru
Python
JavaScript
TypeScript
SQL
Node JS
Apex
Python
Flask
FastAPI
Django
Apex
MuleSoft
Databases
PostgreSQL
Redis
Pinecone
FAISS
RabbitMQ
Apache Kafka
Azure Cosmos DB
AI/ML
LangChain
Vertex AI
Embeddings
Prompt Engineering
AI Agents
Semantic Kernel
AWS Bedrock
RAG
OpenAI
Frontend
GraphQL
Angular
React.js
DevOps
Rest API
Terraform
GCP
OpenShift
Azure DevOps
GitHub Actions
Kong
Azure
CI/CD
GitOps
ArgoCD
Jenkins
AWS
Docker
Kubernetes
Platform Engineering
Apply
≈ $14k – $37k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • Bengaluru
Java
Java
Apache Tomcat
AI/ML
Machine Learning
DevOps
CI/CD
AWS
Kubernetes
Amazon EC2
Amazon ECS
Amazon Kinesis
Linux
Cybersecurity
Okta
Apply
≈ $21k – $48k per year (Estimated) • Hybrid • 5+ years exp • Bachelor's Degree • Bengaluru
Python
Go
SQL
Databases
MySQL
PostgreSQL
Redis
AI/ML
Machine Learning
DevOps
Terraform
Ansible
GCP
Helm
Loki
OpenTelemetry
CloudFormation
Prometheus
GitLab CI
CI/CD
GitOps
ArgoCD
AWS
Kubernetes
Grafana
Spinnaker
Configuration Management
Karpenter
Amazon EKS
Google GKE
IAM
Amazon ECS
Linux
DNS
Cybersecurity
Okta
HashiCorp Vault
Apply
≈ $21k – $48k per year (Estimated) • In office • Bengaluru
Python
Go
Databases
MySQL
PostgreSQL
Redis
OpenSearch
AI/ML
Machine Learning
DevOps
Terraform
GCP
Helm
CI/CD
GitOps
AWS
Kubernetes
Platform Engineering
DNS
Cybersecurity
Okta
Apply
≈ $22k – $62k per year (Estimated) • Hybrid • 13+ years exp • Bachelor's Degree • Hyderabad
Python
Go
Databases
Snowflake
Apache Kafka
Amazon Redshift
DevOps
Terraform
GitHub Actions
OpenTelemetry
Terragrunt
Datadog
CI/CD
ArgoCD
AWS
Kubernetes
Platform Engineering
Service Mesh
AWS Lambda
Amazon EC2
Amazon S3
IAM
Amazon Kinesis
Cybersecurity
SOC 2
HIPAA
Zero Trust
Apply
≈ $62k – $169k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Beirut
JavaScript
TypeScript
SQL
C#
Node JS
AI/ML
Claude Code
Prompt Engineering
AI Agents
Agentic Workflows
DevOps
Rest API
Terraform
Datadog
CI/CD
AWS
AWS Lambda
GitLab
Management
Linear
Notion
Kanban
Apply
≈ $34k – $84k per year (Estimated) • Remote (likely EAEU) • 3+ years exp • Moscow
AI/ML
ClearML
DevOps
Terraform
Helm
Loki
Prometheus
CI/CD
ArgoCD
Kubernetes
Grafana
Linux
Apply
≈ $14k – $37k per year (Estimated) • In office • 2+ years exp • Mumbai
JavaScript
TypeScript
SQL
C#
C#
ASP.NET Core
Databases
PostgreSQL
MS SQL
Frontend
Redux
React.js
DevOps
GitHub Actions
OpenTelemetry
CI/CD
Git
Docker
Apply
≈ $17k – $37k per year (Estimated) • In office • 6+ years exp • Chennai
Python
JavaScript
Node JS
Bash
DevOps
Splunk
Terraform
Ansible
Red Hat
OpenShift
Podman
VMWare
Dynatrace
Prometheus
CI/CD
Docker
Kubernetes
Grafana
Platform Engineering
Tekton
Linux
Apply
≈ $18k – $41k per year (Estimated) • In office • 5+ years exp • Bengaluru
Python
AI/ML
Anomaly Detection
DevOps
Splunk
Terraform
GCP
Azure DevOps
GitHub Actions
Loki
New Relic
OpenTelemetry
OpenTofu
Datadog
Dynatrace
PagerDuty
Prometheus
GitLab CI
Azure
CI/CD
Jenkins
AWS
Kubernetes
Grafana
Platform Engineering
Configuration Management
AppDynamics
Opsgenie
AIOps
Incident Management
IAM
Management
Jira
ServiceNow
ITSM
Apply
≈ $9.5k – $25k per year (Estimated) • In office • 5+ years exp • Bengaluru
PowerShell
AI/ML
Copilot
Copilot Studio
Cybersecurity
Microsoft Entra ID
DLP
Management
ServiceNow
Power Automate
Power Apps
Microsoft Teams
OneDrive
SharePoint
ITIL
ITSM
Apply
Staff Cloud Engineer 3 months ago
≈ $22k – $52k per year (Estimated) • In office • 5+ years exp • Bengaluru
Python
Go
JavaScript
PowerShell
Node JS
Bash
AI/ML
AI Agents
LLM Guardrails
Agentic Workflows
DevOps
Terraform
Ansible
GCP
GitHub Actions
Terragrunt
GitLab CI
Azure
CI/CD
Jenkins
AWS
Kubernetes
Platform Engineering
IAM
Cybersecurity
Open Policy Agent
SIEM
Apply
≈ $20k – $47k per year (Estimated) • Hybrid • 10+ years exp • Bachelor's Degree • Bengaluru
Java
Databases
Oracle
DevOps
Rest API
Linux
SOAP
Analytics
Oracle Data Integrator
Master Data Management
Apply
≈ $14k – $32k per year (Estimated) • In office • 6+ years exp • Bachelor's Degree • Bengaluru
Java
Databases
Oracle
DevOps
Rest API
Linux
SOAP
Analytics
Oracle Data Integrator
Master Data Management
Apply
≈ $16k – $41k per year (Estimated) • In office • Full-Time • 5+ years exp • Chennai
Java
AI/ML
Copilot
Claude Code
AI Agents
Mobile
JUnit
DevOps
Rest API
New Relic
Jenkins
Git
AWS
Cybersecurity
Sumo Logic
QA
TestNG
Selenium
Playwright
Apply
Hybrid • Full-Time • Bachelor's Degree • Chennai
AI/ML
AI Agents
Edge AI
Management
Microsoft Office
Apply
≈ $13k – $31k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Chennai
JavaScript
SQL
C#
Apex
C#
.NET
Apex
Lightning Web Components
Gearset
Copado
MuleSoft
AI/ML
Copilot
Cursor
Claude Code
AI Agents
Agentforce
DevOps
Rest API
CI/CD
SOAP
Management
Monday.com
Agile
Scrum
Waterfall
QA
Jest
Apply
≈ $16k – $42k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Chennai
Apply
AI Enablement Lead 6 hours ago
≈ $22k – $48k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Chennai
Apply
See all jobs
This is one of many
1,011,843 more open roles from verified company boards, updated every day.