368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$39k – $97k per year (Estimated)
Location
Remote/Hybrid (Guadalajara, Mexico)
Seniority
Senior
Overview
Company
Impact
Profile match
Brillio is a global digital technology consulting and services firm headquartered in Edison, New Jersey, and founded in 2014. The company provides a range of services including artificial intelligence engineering, data analytics, cloud infrastructure management, and product design to help enterprises modernize their digital operations. Operating across North America, Europe, and Asia, the firm serves clients in sectors such as financial services, healthcare, and retail through a network of global delivery centers and strategic technology partnerships.

Senior Observability / Monitoring Engineer

Role Overview

  • We are looking for a Senior Observability / Monitoring Engineer to design, implement, and optimize observability solutions for large-scale enterprise platforms. This role will play a critical part in enabling proactive monitoring, faster incident detection, and improved system reliability across Salesforce and Microsoft Azure environments. The ideal candidate will have strong expertise in metrics, logs, traces, alerting strategies, and observability tooling, along with hands-on experience supporting production environments.

Key Responsibilities - Observability Engineering:

  • Design and implement end-to-end observability frameworks across Salesforce and Azure platforms
  • Establish unified monitoring across logs, metrics, and distributed tracing
  • Define and standardize observability best practices, dashboards, and alerting strategies
  • Enable proactive detection of issues through intelligent alerting and anomaly detection Monitoring & Tooling
  • Implement and manage tools such as Azure Monitor, Application Insights, Splunk, Datadog, Grafana, Prometheus, or similar
  • Build actionable dashboards for operations, SRE, and business stakeholders
  • Optimize alert noise reduction and improve signal-to-noise ratio
  • Continuously enhance monitoring coverage across applications and infrastructure Incident Support & Reliability
  • Support incident management by providing deep insights using observability data
  • Perform root cause analysis (RCA) leveraging logs, traces, and metrics
  • Collaborate with SRE and engineering teams to improve system reliability and performance
  • Contribute to post-incident reviews and continuous improvement initiatives Automation & Integration
  • Automate monitoring setup and configuration using Infrastructure as Code (IaC)
  • Integrate observability tools with CI/CD pipelines and DevOps workflows
  • Develop scripts or tools to enhance monitoring capabilities and data collection Platform & Integration Support
  • Monitor and optimize Salesforce applications, including integrations and APIs
  • Support Azure-based services, ensuring visibility across compute, storage, and networking layers
  • Ensure end-to-end observability across integrated systems and middleware Governance & Compliance
  • Ensure observability practices align with security and compliance requirements (e.g., SOX)
  • Maintain documentation, runbooks, and monitoring standards
  • Support audits and governance reviews as required

Required Skills & Qualifications Technical Skills

  • Strong experience in observability, monitoring, or SRE roles
  • Hands-on experience with Azure (Azure Monitor, Application Insights)
  • Experience with observability tools (Splunk, Datadog, Prometheus, Grafana, etc.)
  • Strong understanding of logs, metrics, traces, and distributed systems
  • Experience with APM tools and performance tuning
  • Scripting skills (Python, PowerShell, Bash, or similar)
  • Familiarity with CI/CD tools (Azure DevOps, Jenkins, GitHub Actions)
  • Knowledge of Infrastructure as Code (Terraform, ARM, Bicep) Platform Knowledge
  • Experience supporting Salesforce environments (monitoring integrations, APIs, performance)
  • Understanding of cloud-native and microservices architectures Operational Excellence
  • Experience in incident management and RCA
  • Ability to analyze system performance and recommend improvements
  • Strong troubleshooting and analytical skills Soft Skills
  • Strong communication and collaboration skills
  • Ability to work with cross-functional and global teams
  • Proactive mindset with a focus on continuous improvement
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Guadalajara
$19k – $47k per year (Estimated) • Remote • Full-Time • Perm
C#
C#
ASP.NET Core
Dapper
Entity Framework Core
Databases
Apache Kafka
ClickHouse
ElasticSearch
PostgreSQL
RabbitMQ
Redis
DevOps
CI/CD
Docker
Docker Compose
GitHub Actions
Kubernetes
TeamCity
GitHub
GitLab
Apply
$17k – $43k per year (Estimated) • In office • Full-Time • Tomsk
C#
Java
Python
DevOps
CI/CD
Docker
Git
Grafana
Graylog
Kubernetes
Prometheus
Splunk
Apply
$19k – $47k per year (Estimated) • Remote • Full-Time • Tomsk
C#
C#
ASP.NET Core
Dapper
Entity Framework Core
Databases
Apache Kafka
ClickHouse
ElasticSearch
PostgreSQL
RabbitMQ
Redis
DevOps
CI/CD
Docker
Docker Compose
GitHub Actions
Kubernetes
TeamCity
GitHub
GitLab
Apply
$17k – $43k per year (Estimated) • In office • Full-Time • Perm
C#
Java
Python
DevOps
CI/CD
Docker
Git
Grafana
Graylog
Kubernetes
Prometheus
Splunk
Apply
$18k – $48k per year (Estimated) • Remote/Hybrid • Full-Time • Moscow
Python
C++
C++
CMake
DevOps
Bitbucket
CI/CD
Git
Gitflow
Jenkins
RTOS
Chips/EDA
OpenOCD
Management
Confluence
Jira
QA
Pytest
Apply
$150k – $160k per year • In office • 8+ years exp • Bachelor's Degree • New York
Java
Node JS
Python
TypeScript
JavaScript
Databases
DynamoDB
Frontend
React.js
DevOps
Amazon EC2
AWS
AWS Lambda
CI/CD
CloudFormation
Kubernetes
Terraform
Amazon S3
IAM
Apply
$23k – $57k per year (Estimated) • Remote/Hybrid • 7+ years exp • Bengaluru
SQL
Databases
Azure SQL Database
Oracle
PostgreSQL
DevOps
Azure
GCP
Incident Management
Apply
$32k – $60k per year (Estimated) • In office • 4+ years exp • Bachelor's Degree • Mumbai
Design
Adobe XD
FigJam
Figma
Sketch
Management
Miro
Apply
$22k – $62k per year (Estimated) • In office • 4+ years exp • Bengaluru
JavaScript
Python
DevOps
CI/CD
QA
Cypress
Playwright
Postman
Rest-Assured
Selenium
Apply
$24k – $64k per year (Estimated) • In office • 7+ years exp • Bachelor's Degree • Bengaluru
Java
Python
TypeScript
Apex
Apex
MuleSoft
AI/ML
AI Agents
LangGraph
Semantic Kernel
LangChain
Function Calling
Human-in-the-Loop
Model Context Protocol
OpenAI
DevOps
AWS
Azure
Kubernetes
Apply
$20k – $53k per year (Estimated) • In office • Full-Time • 1+ year exp • High School Diploma • Guadalajara
Cybersecurity
HIPAA
Apply
$23k – $53k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Guadalajara
SQL
Apply
$58k – $139k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Guadalajara
Python
AI/ML
Knowledge Graph
Kubeflow
Apply
$28k – $94k per year (Estimated) • Remote • Full-Time • 4+ years exp • Guadalajara
PHP
JavaScript
PHP
Eloquent
Laravel
Livewire
Pest
PHPUnit
Frontend
Inertia.js
React.js
Tailwind CSS
Vue.js
QA
Playwright
Apply
$41k – $117k per year (Estimated) • Remote • Full-Time • 5+ years exp • Guadalajara
JavaScript
Python
TypeScript
Python
Django
FastAPI
Flask
Databases
PostgreSQL
Frontend
React.js
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.