{"id":1225571,"url":"https://alion.io/job/ampera-technologies-senior-ai-platform-mlops-engineer","title":"Senior AI Platform / MLOps Engineer","company":{"id":3798322,"name":"Ampera Technologies","domain":"amperatech.ai","url":"https://alion.io/company/ampera-technologies","size_band":"5000+","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":null,"truth_index":null},"role":"AI/ML","role_family":"AI/ML","seniority":"senior","employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"board_field","remote_working_hours":null,"hiring_geo_confidence":"inferred","locations":["Chennai, India","Bengaluru, India","Mumbai, India","Delhi, India","Gurgaon, India","Noida, India","Ghaziabad, India","Faridabad, India","Pune, India","Hyderabad, India","Kolkata, India"],"countries":["IN"],"hiring_countries":["IN"],"hiring_countries_total":1,"salary":null,"salary_estimate":{"min_usd":40000,"max_usd":101000,"period":"year","method":"role_seniority_country_cell","sample_n":10},"experience_years_min":6,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Amazon S3","optional":false},{"name":"ArgoCD","optional":false},{"name":"AWS","optional":false},{"name":"CI/CD","optional":false},{"name":"Grafana","optional":false},{"name":"KServe","optional":false},{"name":"Kubernetes","optional":false},{"name":"KV Cache","optional":false},{"name":"LLM","optional":false},{"name":"Milvus","optional":false},{"name":"OpenShift","optional":false},{"name":"pgvector","optional":false},{"name":"Platform Engineering","optional":false},{"name":"Prometheus","optional":false},{"name":"Splunk","optional":false},{"name":"TensorRT-LLM","optional":false},{"name":"Triton","optional":false},{"name":"vLLM","optional":false},{"name":"PostgreSQL","optional":true},{"name":"TensorRT","optional":true}],"status":"live","first_seen_at":"2026-09-24T10:27:44Z","employer_posted_date":null,"last_verified_at":"2026-09-24T10:27:44Z","board_verified":false,"closed_at":null,"days_open":2,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":2},"description":"Title : Senior AI Platform / MLOps Engineer\nExperience : 6+ years\nWork type : Chennai - Work from Office/other locations - Remote\nEmployment Type : Full Time\nNotice Period : Immediate\nWork Day :Mon to Fri\n\nKey Responsibilities:\nInstall, configure and operate OpenShift, NVIDIA GPU operator, OpenShift AI, and NIM microservices on 12× RTX PRO 6000 across two servers; single-node and HA control-plane topologies\nServing configuration and tuning: quantized model deployment (FP8/FP4), replica balancing, batching, KV-cache and context management\nAzure GPU build environments: provisioning, cost control, parity with the on-prem stack via pinned container/model versions; cloud-to-factory migration with parity regression\nGitOps CI/CD, container registry, artifact/model versioning, environment promotion; observability and audit wiring (Splunk, Prometheus/Grafana)\nBenchmark automation: load harness, p50/p95/p99 latency, tokens/sec, GPU utilization; the capacity report data pipeline\nPlatform upgrade procedure with evaluation-regression gates; deployment runbook as a first-class deliverable\nTechnical Skills:\n6+ years infrastructure/platform engineering with 3+ years production Kubernetes; OpenShift experience strongly preferred\nHands-on GPU inference serving in production: NIM, Triton, vLLM, or TensorRT-LLM — you have sized, deployed, and tuned LLM serving on real GPUs and can talk memory-bandwidth trade-offs\nGitOps fluency (ArgoCD/Flux), infrastructure-as-code, container internals; comfortable in air-gapped/proxy-restricted enterprise networks\nObservability depth: metrics, traces, log pipelines; has built performance test harnesses, not just run them\n\nAzure or AWS GPU compute operations experience\nStrongly preferred\nNVIDIA GPU operator and AI Enterprise stack specifics; KServe; Milvus or pgvector operations; VAST/NFS/S3 storage integration; banking or other regulated-environment delivery\n\nAbout Ampera: \nAmpera Technologies, a purpose driven Digital IT Services with primary focus on supporting our client with their Data, AI / ML, Accessibility and other Digital IT needs. We also ensure that equal opportunities are provided to Persons with Disabilities Talent. Ampera Technologies has its Global Headquarters in Chicago, USA and its Global Delivery Center is based out of Chennai, India. We are actively expanding our Tech Delivery team in Chennai and across India. We offer exciting benefits for our teams, such as 1) Hybrid and Remote work options available, 2) Opportunity to work directly with our Global Enterprise Clients, 3) Opportunity to learn and implement evolving Technologies, 4) Comprehensive healthcare, and 5) Conducive environment for Persons with Disability Talent meeting Physical and Digital Accessibility standards \nSkills\nMachine Learning (ML), MLOps, Kubernetes, Large Language Models (LLM), openshift, CI/CD","description_format":"text","description_chars":2852,"description_truncated":false,"requirements":{"experience_years_min":6,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[{"name":"India","iso":"IN","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Decision Intelligence","AI Consulting & Integration","Analytics & BI Consulting"],"lifecycle":[{"event":"open","at":"2026-09-25T13:06:44Z"}],"liveness":{"score":86,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.86,"p_room":1,"age_days":1,"expected_fill_days":20,"reasons":["seen:1","win:early"],"computed_at":"2026-09-26T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/ampera-technologies-senior-ai-platform-mlops-engineer","json_url":"https://alion.io/job/ampera-technologies-senior-ai-platform-mlops-engineer.json","meta":{"generated_at":"2026-09-27T03:49:14Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":3558,"day_limit":5000,"remaining_today":1442,"minute_limit":60,"resets_at":"2026-09-28T00:00:00Z"}}}