Confirmed on the employer's own hiring board on Oct 9, 2026. First seen by Alion on Oct 8, 2026.
Job Summary
10+ years of experience in DevOps, Site Reliability Engineering (SRE), and Production Support environments (SaaS product support)
Strong expereince in driving and leading team of 10+ members
Strong hands-on expertise in Core Java, Unix/Linux administration, Shell Scripting(python/bash), and automation.
Experience handling 24x7 on-call support, incident management, service restoration, and customer escalations.
Proven ability to perform root cause analysis (RCA), troubleshooting, and resolution of complex production issues.
Good understanding of distributed systems and distributed database architectures.
Hands-on experience in installation, upgrade, maintenance, backup, and performance tuning of PostgreSQL, ZooKeeper (ZK), and OpenLDAP.
Experience with Cassandra database administration and support is a strong plus.
Knowledge of monitoring, observability, log analysis, and system health management tools.
Strong ownership mindset with the ability to drive issues independently and collaborate with cross-functional teams.
Excellent verbal and written communication skills with the flexibility to work in a fast-paced, global support environment.
Key Responsibilities
10+ years of experience in DevOps, Site Reliability Engineering (SRE), and Production Support environments (SaaS product support)
Strong expereince in driving and leading team of 10+ members
Strong hands-on expertise in Core Java, Unix/Linux administration, Shell Scripting(python/bash), and automation.
Experience handling 24x7 on-call support, incident management, service restoration, and customer escalations.
Proven ability to perform root cause analysis (RCA), troubleshooting, and resolution of complex production issues.
Good understanding of distributed systems and distributed database architectures.
Hands-on experience in installation, upgrade, maintenance, backup, and performance tuning of PostgreSQL, ZooKeeper (ZK), and OpenLDAP.
Experience with Cassandra database administration and support is a strong plus.
Knowledge of monitoring, observability, log analysis, and system health management tools.
Strong ownership mindset with the ability to drive issues independently and collaborate with cross-functional teams.
Excellent verbal and written communication skills with the flexibility to work in a fast-paced, global support environment.
Skill Requirements
10+ years of experience in DevOps, Site Reliability Engineering (SRE), and Production Support environments (SaaS product support)
Strong expereince in driving and leading team of 10+ members
Strong hands-on expertise in Core Java, Unix/Linux administration, Shell Scripting(python/bash), and automation.
Experience handling 24x7 on-call support, incident management, service restoration, and customer escalations.
Proven ability to perform root cause analysis (RCA), troubleshooting, and resolution of complex production issues.
Good understanding of distributed systems and distributed database architectures.
Hands-on experience in installation, upgrade, maintenance, backup, and performance tuning of PostgreSQL, ZooKeeper (ZK), and OpenLDAP.
Experience with Cassandra database administration and support is a strong plus.
Knowledge of monitoring, observability, log analysis, and system health management tools.
Strong ownership mindset with the ability to drive issues independently and collaborate with cross-functional teams.
Excellent verbal and written communication skills with the flexibility to work in a fast-paced, global support environment.

