368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$23k – $58k per year (Estimated)
Location
In office (Pune)
Seniority
Senior · 6+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Qualys, Inc. is a publicly traded enterprise cybersecurity, compliance, and IT solutions provider headquartered in Foster City, California. Founded in 1999 as a pioneer in SaaS-delivered vulnerability management, the company provides the Qualys Enterprise TruRisk Platform. Its single-agent cloud architecture unifies cybersecurity asset management, vulnerability management detection and response (VMDR), patch management, web application security, policy compliance, and cloud security across hybrid IT environments, public clouds, and containers.

Come work at a place where innovation and teamwork come together to support the most exciting missions in the world!

Role Overview

We are seeking a Senior DevOps Engineer, Observability to design, build, operate, and continuously improve a high-scale observability platform built around ClickHouse, HyperDX, OpenTelemetry, Kubernetes, and modern DevOps automation practices. This role is intended for a senior hands-on engineer who can take end-to-end ownership of observability infrastructure for production environments. The engineer will be responsible for building reliable telemetry pipelines for logs, metrics, and distributed traces, optimizing ClickHouse for large-scale observability workloads, operating HyperDX for troubleshooting and application performance analysis, and partnering with application, platform, and SRE teams to improve production visibility, reliability, and incident response. The ideal candidate is deeply technical, operationally disciplined, automation-oriented, and comfortable working in high-volume, production-critical environments. Senior-level expectation: This role requires ownership beyond task execution, including technical judgment, production accountability, automation-first delivery, mentoring, and clear communication during incidents and escalations.

Key Responsibilities

  • Design, build, and operate scalable observability platforms using ClickHouse, HyperDX, OpenTelemetry, Kubernetes, Prometheus, Grafana, Alertmanager, Fluent Bit, and Filebeat.
  • Architect and optimize ClickHouse for high-volume observability workloads, including logs, traces, metrics, and telemetry analytics.
  • Design and manage ClickHouse schemas, partitioning strategies, ordering keys, TTLs, materialized views, retention policies, storage efficiency, and query optimization.
  • Build, operate, and improve HyperDX for log search, distributed tracing, service analysis, dashboards, telemetry correlation, troubleshooting, and root-cause analysis.
  • Build and maintain scalable OpenTelemetry Collector pipelines for collecting, processing, enriching, filtering, sampling, and routing telemetry data.
  • Implement reliable correlation across logs, metrics, and traces to support faster application troubleshooting, service dependency analysis, and incident resolution.
  • Design resilient telemetry pipelines with batching, queuing, retries, backpressure handling, sampling, rate limiting, and cardinality controls.
  • Deploy and operate observability infrastructure on Kubernetes, with focus on scalability, high availability, capacity planning, resiliency, and operational safety.
  • Automate infrastructure deployment, configuration management, platform upgrades, application onboarding, and recurring operational tasks using DevOps best practices.
  • Build and maintain CI/CD workflows using Jenkins, infrastructure automation using Terraform and Ansible, and service discovery or secrets management integrations using HashiCorp Consul and Vault.
  • Partner with application engineering, platform engineering, SRE, and operations teams to troubleshoot production issues, improve observability coverage, and reduce mean time to detect and resolve incidents.
  • Analyze production performance issues across infrastructure, applications, telemetry pipelines, and ClickHouse queries, and drive corrective actions to closure.
  • Define and improve standards for telemetry instrumentation, log quality, metric hygiene, trace propagation, dashboard design, alert quality, and production readiness.
  • Participate in incident response, post-incident reviews, capacity planning, operational reviews, and remediation tracking for observability services.
  • Mentor junior engineers, review designs and automation changes, and raise the overall technical and operational maturity of the team. Required Qualifications
  • 6+ years of experience in DevOps, SRE, Platform Engineering, Infrastructure Engineering, or Observability Engineering roles.
  • Strong production experience with ClickHouse, including architecture, administration, schema design, performance tuning, query optimization, troubleshooting, retention management, and operations at scale.
  • Hands-on experience deploying and operating HyperDX, including integration with ClickHouse and usage for log search, distributed tracing, dashboards, service analysis, and troubleshooting.
  • Strong experience with OpenTelemetry and OpenTelemetry Collector, including receiver, processor, exporter, sampling, enrichment, batching, and routing configurations.
  • Strong working knowledge of Fluent Bit, Filebeat, Prometheus, Alertmanager, Grafana, ClickHouse, HyperDX, and OpenTelemetry.
  • Strong DevOps and automation experience with Jenkins, CI/CD pipelines, Ansible, Terraform, HashiCorp Consul, HashiCorp Vault, and Kubernetes.
  • Strong understanding of Linux, networking, microservices, REST, gRPC, distributed systems, high-availability architectures, and production operations.
  • Experience operating and troubleshooting high-volume production systems with focus on reliability, scalability, capacity, and performance.
  • Ability to debug complex issues across application telemetry, Kubernetes infrastructure, data ingestion pipelines, storage systems, and query performance.
  • Strong scripting and automation skills, with the ability to reduce manual operations and improve repeatability, reliability, and operational efficiency.
  • Ability to take ownership of production systems, drive problems to closure, communicate clearly during incidents, and collaborate effectively with engineering and operations teams. Preferred Qualifications
  • Development experience with Java and Python, including the ability to understand application code, troubleshoot instrumentation issues, and analyze performance behavior.
  • Experience developing automation, internal tools, APIs, platform services, or self-service onboarding workflows.
  • Experience with Kafka or other high-throughput messaging and streaming platforms.
  • Strong understanding of APM, distributed tracing, OpenTelemetry instrumentation, context propagation, service maps, and service dependency analysis.
  • Experience with cloud platforms such as AWS, Azure, GCP, or OCI.
  • Experience designing observability solutions for large-scale Kubernetes, microservices, or distributed application environments.
  • Experience improving alert quality, reducing noise, defining SLOs or SLIs, and supporting production readiness reviews. Senior Engineer Expectations As a Senior DevOps Engineer, this role is expected to go beyond task execution and operate with strong ownership, judgment, technical leadership, and accountability for production outcomes. The candidate should be able to:
  • Own critical observability services end to end, from design and implementation to production operations, troubleshooting, and continuous improvement.
  • Make sound technical decisions around scalability, reliability, performance, cost, capacity, security, and operational maintainability.
  • Proactively identify risks, gaps, and scaling bottlenecks before they become production issues.
  • Build automation-first solutions rather than relying on manual processes.
  • Lead technical investigations during complex incidents and drive clear remediation plans.
  • Influence application and platform teams to improve instrumentation quality, telemetry consistency, and operational readiness.
  • Write clear technical documentation, operational runbooks, onboarding guides, and incident review notes.
  • Mentor engineers, review technical designs, and help establish standards for observability engineering.
  • Communicate effectively with technical and non-technical stakeholders during incidents, escalations, and planning discussions.
  • Demonstrate strong accountability for production stability, platform reliability, and measurable operational outcomes.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Pune
$16k – $38k per year (Estimated) • In office • Full-Time • 3+ years exp • India
SQL
Databases
Azure SQL Database
DevOps
AWS
Azure
CI/CD
GitHub
Kubernetes
Terraform
Apply
$24k – $63k per year (Estimated) • In office • Full-Time • 6+ years exp • Master's Degree • India
Crystal
Groovy
JavaScript
Perl
Python
Ruby
SQL
TypeScript
Java
Java
Apache Tomcat
Gradle
Hibernate
Maven
Spring Boot
Spring MVC
Databases
Apache Kafka
Db2
Oracle
PostgreSQL
RabbitMQ
AI/ML
Fine-tuning
Frontend
Angular
JQuery
DevOps
Apache HTTP Server
AWS
Azure
CI/CD
Docker
GCP
Jenkins
Kubernetes
Rest API
Cybersecurity
Checkmarx
SonarQube
Apply
$96k – $163k per year • Equity • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Boston
C#
JavaScript
Python
TypeScript
DevOps
Azure AKS
CI/CD
Kubernetes
Rest API
Azure
Apply
$64k – $189k per year (Estimated) • Remote/Hybrid • Full-Time • 1+ year exp • Bachelor's Degree • Singapore
Python
SQL
Databases
Apache Kafka
AI/ML
Amazon SageMaker
Kubeflow
MLFlow
Spark
Vertex AI
DevOps
AWS
Azure
Azure DevOps
CI/CD
Docker
GCP
GitLab
GitLab CI
Jenkins
Kubernetes
Apply
$142k – $213k per year • Remote/Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Jersey City
Java
Python
SQL
TypeScript
JavaScript
Java
Spring Boot
Python
Asyncio
FastAPI
AI/ML
AI Agents
Claude
Claude Code
Copilot
Cursor
Devin
Fine-tuning
Gemini
Google ADK
Hybrid Search
Knowledge Graph
LangChain
LangGraph
RAG
Frontend
Angular
DevOps
CI/CD
Docker
Kubernetes
Rest API
Apply
Lead SFDC Developer 5 days ago
$155k – $175k per year • Equity • In office • Full-Time • 7+ years exp • Bachelor's Degree • Foster City
Apex
Apex
Lightning Web Components
Visualforce
DevOps
Bitbucket
CI/CD
Git
Cybersecurity
Okta
Qualys Cloud Platform
Marketing
Salesforce
Apply
$215k – $255k per year • Equity • In office • Full-Time • 16+ years exp • Bachelor's Degree • Foster City
DevOps
AWS
Azure
GCP
Platform Engineering
Cybersecurity
ISO 27001
Qualys Cloud Platform
Apply
$31k – $68k per year (Estimated) • In office • Full-Time • 10+ years exp • Pune
AI/ML
AI Agents
Cybersecurity
Qualys Cloud Platform
Design
Figma
Management
Jira
ServiceNow
Apply
$136k – $212k per year (Estimated) • In office • Full-Time • 5+ years exp • Georgia
Cybersecurity
HIPAA
ISO 27001
Qualys Cloud Platform
Apply
$18k – $59k per year (Estimated) • In office • Full-Time • 5+ years exp • Pune
JavaScript
Node JS
TypeScript
AI/ML
Claude
Copilot
LLM
Prompt Engineering
AI Agents
Devin
Model Context Protocol
Frontend
React.js
Redux
Redux Toolkit
Vite
Webpack
Mobile
State Management
DevOps
Bitbucket
CI/CD
Docker
Git
Jenkins
Kubernetes
Rest API
Design
Figma
Management
Jira
QA
Mocha
Vitest
Apply
$12k – $27k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Pune
SQL
Analytics
Power BI
Tableau
Apply
$13k – $29k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Pune
JavaScript
Apex
Apex
MuleSoft
AI/ML
AI Agents
Edge AI
DevOps
AWS
Azure
Management
Draw.io
Marketing
Salesforce
Apply
$11k – $42k per year (Estimated) • In office • Full-Time • 4+ years exp • Bachelor's Degree • Pune
ABAP
Apply
Data Architect 4 hours ago
$38k – $91k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru • Pune
Node JS
Python
SQL
JavaScript
Databases
Databricks
MongoDB
Redis
Apply
$23k – $62k per year (Estimated) • In office • Full-Time • 3+ years exp • Navi Mumbai • Pune
Python
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.