{"id":1265936,"url":"https://alion.io/job/coredge-io-senior-devops-engineer","title":"Senior DevOps Engineer","company":{"id":3801269,"name":"Coredge.io","domain":"coredge.io","url":"https://alion.io/company/coredge-io","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":null,"truth_index":null},"role":"DevOps","role_family":"DevOps","seniority":"senior","employment_type":null,"work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Noida, India"],"countries":["IN"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":16000,"max_usd":36000,"period":"year","method":"role_seniority_country_remote_cell","sample_n":42},"experience_years_min":4,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Ansible","optional":false},{"name":"Argo Workflows","optional":false},{"name":"ArgoCD","optional":false},{"name":"Grafana","optional":false},{"name":"Kubernetes","optional":false},{"name":"Linux","optional":false},{"name":"Prometheus","optional":false},{"name":"Python","optional":false}],"status":"live","first_seen_at":"2026-09-08T11:27:41Z","employer_posted_date":null,"last_verified_at":"2026-09-08T11:27:41Z","board_verified":false,"closed_at":null,"days_open":20,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":20},"description":"About the Role :\n\nWe are seeking a highly skilled and motivated Senior DevOps Engineer to join our team. The ideal candidate will have 4 - 5 years of DevOps experience with a strong focus on on-premises or self-managed Kubernetes environments.\n\nThis role involves deploying, operating, monitoring, and troubleshooting Kubernetes clusters and Linux infrastructure to ensure high availability, reliability, and performance of production systems.\n\nKey Requirements :\n\n- Bachelor's degree in computer science, IT, or a related field (or equivalent hands-on experience).\n\n- 4 - 5 years of experience in a DevOps / SRE / Production Support role.\n\n- Strong expertise in Linux system administration.\n\n- Hands-on experience with on-prem, self-managed, or unmanaged Kubernetes clusters.\n\n- Proven ability to deploy, debug, and troubleshoot Kubernetes environments.\n\n- Strong experience with Prometheus and Grafana.\n\n- Mandatory exposure to ARGO CD / ARGO Workflows.\n\n- Experience with automation and scripting (Shell, Python, Ansible).\n\n- Ability to handle production incidents independently.\n\n- Excellent troubleshooting, analytical, and communication skills.\n\nKey Responsibilities :\n\n1. Linux Administration :\n\n- Administer and support Linux-based systems in production environments.\n\n- Deploy, manage, and troubleshoot applications running on Linux.\n\n- Perform root cause analysis for OS-level issues to maintain high availability and performance.\n\n- Ensure system stability, security hardening, and performance tuning.\n\n2. Kubernetes (On-Prem / Self-Managed) - Must Have :\n\n- Deploy, configure, and maintain on-premises or self-managed Kubernetes clusters (bare metal or VM-based).\n\n- Troubleshoot Kubernetes issues related to pod scheduling, networking, storage, cluster components, and service failures.\n\n- Debug containerized workloads and ensure reliable rollout.\n\n3. Monitoring & Observability - Must Have :\n\n- Implement and maintain Prometheus and Grafana for infrastructure and application monitoring.\n\n- Create and manage real-time Grafana dashboards for cluster health, application metrics, and alerts.\n\n- Analyze monitoring data to proactively identify and resolve performance and reliability issues.\n\n4. Automation & CronJobs :\n\n- Configure and manage Kubernetes CronJobs and Linux-based scheduled tasks.\n\n- Improve operational efficiency through scripting and automation using Shell, Python, or Ansible.\n\n5. Platform / Portal Exposure (Good to Have) :\n\n- Gain working knowledge of Horizon / platform portals used for infrastructure or operational visibility.\n\n6. Cloud Awareness (Limited / Supporting Role) : \n\n- Understand basic cloud computing concepts and architectures.\n\n- Note: This role is primarily focused on on-prem Kubernetes, not public cloud operations.\n\nSkills\nDevOps, Kubernetes, Linux, Prometheus, Grafana, Ansible, Python, System Administration, Site Reliability, Production Support, Monitoring Tools","description_format":"text","description_chars":2922,"description_truncated":false,"requirements":{"experience_years_min":4,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Cloud Platforms (IaaS & PaaS)","AI Compute & Inference"],"lifecycle":[{"event":"open","at":"2026-09-25T22:00:00Z"}],"liveness":{"score":44,"band":"fade","label":"Fading","p_open":0.85,"p_active":0.692,"p_room":0.75,"age_days":19,"expected_fill_days":23,"reasons":["seen:19","velocity","win:late"],"computed_at":"2026-09-28T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/coredge-io-senior-devops-engineer","json_url":"https://alion.io/job/coredge-io-senior-devops-engineer.json","meta":{"generated_at":"2026-09-29T02:17:05Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":1922,"day_limit":5000,"remaining_today":3078,"minute_limit":60,"resets_at":"2026-09-30T00:00:00Z"}}}