368,746open jobs
9,444companies
47,506added this week
Browse all
Salary
$23k – $58k per year (Estimated)
Location
In office (Chandigarh)
Seniority
Senior · 6+ years exp
Overview
Company
Impact
Profile match
NeoSOFT is a global IT consulting & software solutions provider specializing in Software Development Services, Mobile Application Development Services, Web Hosting Services, Web Development Services, etc.

We are seeking a Senior Site Reliability Engineer (SRE) to join our engineering team. This is a hands-on individual contributor role for an experienced SRE/DevOps/Platform Engineer who can take ownership of high-severity production incidents and drive continuous improvements in platform reliability and resilience. The successful candidate will work closely with global engineering teams, participate in the on-call rotation, troubleshoot production issues, improve observability, and implement automation to reduce operational toil.

The core responsibilities for the job include the following:

Incident Management and Production Support:

  • Participate in the on-call rotation and lead high-severity production incidents.
  • Drive incident triage, coordinate cross-functional response, and communicate status to stakeholders.
  • Troubleshoot production alerts across AWS infrastructure and application layers.
  • Distinguish transient issues from systemic failures and take appropriate corrective action.

Monitoring and Observability:

  • Own and enhance Datadog monitoring standards.
  • Design and maintain monitors, dashboards, SLOs, and SLIs.
  • Improve signal-to-noise ratios and reduce unnecessary alert fatigue.
  • Monitor system health and identify reliability risks proactively.

Reliability Engineering:

  • Identify systemic weaknesses and implement reliability improvements.
  • Lead initiatives focused on reducing operational toil and inefficiencies.
  • Design and implement automation and process improvements.
  • Drive reliability initiatives through implementation and adoption.

RCA, Runbooks, and Documentation:

  • Conduct detailed Root Cause Analysis (RCA) for infrastructure and application failures.
  • Create structured post-mortems with actionable follow-up items.
  • Develop and maintain operational runbooks and SOPs.
  • Document operational learnings and continuously improve troubleshooting processes.

Deployment and CI/CD:

  • Monitor CI/CD pipelines during deployments.
  • Identify potential reliability risks before and during releases.
  • Initiate rollbacks following established procedures when system stability is at risk.

Collaboration and Mentorship:

  • Mentor junior and mid-level engineers during complex investigations.
  • Review runbooks, incident reports, and operational documentation.
  • Promote strong troubleshooting and operational best practices across the team.
  • Collaborate effectively with global engineering teams across time zones.

Requirements:

  • 6-8 years of hands-on experience in SRE, DevOps, or platform engineering.
  • Proven experience leading responses to high-severity production incidents.
  • Strong hands-on experience with AWS infrastructure.
  • Working knowledge of: Amazon ECS, IAM, VPC, ALB/NLB, RDS, S3 MSK, ElastiCache, Lambda, and CloudWatch.
  • Strong experience with Terraform / Infrastructure as Code.
  • Hands-on experience building and managing Datadog monitors and dashboards.
  • Strong understanding of SLOs, SLIs, and error budgets.
  • Strong Linux command-line skills.
  • Experience writing Python automation scripts.
  • Working knowledge of GitHub Actions or GitLab CI for ECS-based deployments.
  • Understanding of architectural patterns such as microservices, pub/sub, and load balancing.
  • Strong structured troubleshooting and problem-solving skills.
  • Excellent written and verbal English communication skills.

Preferred Skills:

  • Familiarity with Grafana.
  • Experience with RDS or Cassandra performance metrics.
  • Basic understanding of financial markets and market data, including equities, options, and market data feeds.
  • Experience mentoring engineers or establishing operational best practices.
  • Ability to independently learn unfamiliar systems using documentation and runbooks.
  • Strong documentation and knowledge-sharing practices.

The ideal candidate demonstrates:

  • Strong ownership and accountability.
  • A systematic, hypothesis-driven approach to troubleshooting.
  • Ability to remain effective under production pressure.
  • Strong cross-functional communication.
  • Proactive identification and resolution of reliability issues.
  • Ability to drive issues from discovery through complete resolution.
  • A continuous improvement mindset focused on automation, resilience, and operational excellence.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,746 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Chandigarh
$220k – $325k per year • Remote/Hybrid • Full-Time • 15+ years exp • New York
DevOps
AWS
Azure
CI/CD
GCP
Kubernetes
Platform Engineering
Cybersecurity
Threat Modeling
Apply
Remote/Hybrid • 7+ years exp • Bachelor's Degree
C#
JavaScript
Python
SQL
TypeScript
C#
ASP.NET Core
Blazor
Dapper
Databases
Oracle
Frontend
Angular
Bootstrap
JQuery
Vue.js
DevOps
AWS
Azure
Azure DevOps
CI/CD
Git
Apply
$26k – $66k per year (Estimated) • In office • Full-Time • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
$140k – $253k per year (Estimated) • In office • 5+ years exp • Long Beach
C++
Rust
C++
Protobuf
AI/ML
Human-in-the-Loop
DevOps
CI/CD
Vector
Apply
$21k – $57k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Noida
C#
SQL
TypeScript
JavaScript
C#
.NET
Frontend
Angular
DevOps
Azure
Azure DevOps
CI/CD
Git
TeamCity
Apply
Lead Developer (Java) 29 days ago
$29k – $77k per year (Estimated) • In office • Bengaluru
JavaScript
TypeScript
Java
Java
Spring Boot
Frontend
Angular
React.js
Apply
SDET - Java 2 months ago
In office • Bengaluru
Java
SQL
Java
Spring Boot
DevOps
Azure
Azure DevOps
CI/CD
GitHub Actions
Jenkins
Rest API
GitHub
QA
Selenium
Apply
$20k – $83k per year (Estimated) • In office • Mumbai
C#
SQL
JavaScript
C#
ASP.NET Core
Databases
Apache Kafka
Databricks
MS SQL
PostgreSQL
AI/ML
Copilot
Frontend
React.js
DevOps
Azure
Docker
Kubernetes
GitHub
Apply
Java Developer 2 months ago
$24k – $65k per year (Estimated) • In office • 6+ years exp • Bengaluru
Java
TypeScript
JavaScript
Java
Hibernate
Spring Boot
Databases
Apache Kafka
MySQL
Redis
Frontend
Angular
React.js
DevOps
Amazon EC2
AWS
AWS Lambda
CI/CD
Git
Rest API
Amazon S3
Apply
$21k – $91k per year (Estimated) • In office • Bengaluru
JavaScript
TypeScript
Java
Java
Spring Boot
Databases
Apache Kafka
Frontend
Angular
React.js
DevOps
Rest API
Apply
$20k – $50k per year (Estimated) • Remote • Full-Time • 8+ years exp • Chandigarh
C#
Java
Node JS
Python
SQL
TypeScript
JavaScript
C#
.NET
Databases
Apache Kafka
ElasticSearch
MS SQL
PostgreSQL
DevOps
Datadog
Docker
Kubernetes
Nginx
QA
Postman
SoapUI
Apply
$41k – $88k per year (Estimated) • Remote • Full-Time • 10+ years exp • Chandigarh
Bash
Java
PowerShell
Python
SQL
C++
Java
Gradle
Maven
C++
CMake
Databases
MySQL
DevOps
AWS
Azure
Azure DevOps
CI/CD
Docker
GCP
Git
GitHub Actions
GitLab CI
Jenkins
Kubernetes
GitHub
GitLab
Apply
$27k – $67k per year (Estimated) • Remote • Full-Time • 10+ years exp • Chandigarh
C++
Go
JavaScript
Lua
Node JS
SQL
TypeScript
Java
Java
Apache Tomcat
Databases
Apache Kafka
DynamoDB
ElasticSearch
InfluxDB
MySQL
Redis
DevOps
AWS
Docker
Kubernetes
Nginx
Amazon S3
Apply
Remote • Internship • Bachelor's Degree • Chandigarh
DevOps
SLI/SLO/SLA
Analytics
Power BI
Apply
See all jobs
This is one of many
368,746 more open roles from verified company boards, updated every day.