Location
In office (Bengaluru)
Overview
Company
Impact
Profile match
Humynlabs is an AI-first data intelligence platform enabling AI training and evaluation to deliver high-quality multimodal datasets with robust quality control for reliable, production-ready outcomes.
Responsibilities:
- Productionize research prototypes: Take model code from researchers (PyTorch, CUDA, exotic dependency stacks) and deliver production pipeline stages: reproducible Docker images, pinned GPU/CUDA/cuDNN environments, clean I/O contracts, retries, idempotency, and observability.
- A researcher's stereo depth-estimation model: a GPU batch stage with a locked CUDA base image, standardized S3 input/output layout, and per-clip cost tracking.
- A fused MediaPipe + WiLoR hand-pose prototype: a single orchestrated labelling stage with well-defined intermediate artifacts and failure isolation per clip.
- A monocular depth model + a stereo model: a three-stage fuse pipeline where resolution-alignment invariants are enforced by the platform, not by tribal knowledge.
- Own the orchestration layer: Design and evolve Step Functions state machines, AWS Batch compute environments and job queues, Lambda glue, and selective stage re-execution (re-run just one stage across a fleet of clips without redoing everything).
- Own the infrastructure as code: All of it lives in Terraform modules, per-environment stacks, ECR, IAM, networking. You'll extend and harden this, not click around a console.
- Drive cost efficiency as a first-class feature: Spot capacity strategies, right-sizing GPU instance families, eliminating GPU idle time (we've measured it, we hunt it), storage lifecycle policies on multi-TB S3 datasets, and batching strategies that keep expensive GPUs saturated. You should be the person who can say what a pipeline run costs per clip and then make that number go down.
- Make it reliable at scale: structured logging, metrics, and alerting across stages; dead-letter handling and automatic retries for flaky clips; data-quality gates so bad inputs fail fast and loudly instead of silently poisoning downstream datasets.
- Manage the container fleet: A dozen-plus GPU images with heavy, conflicting ML dependencies. Keep builds fast, images slim, CUDA stacks consistent, and breakage (e. g., an upstream wheel disappearing from an index) fixed within hours, not weeks.
- Move fast with researchers: Sit close to the research loop prototype, deploy to staging, run on real fleet data, iterate on feedback, and promote to production.
Requirements:
- 5+ years in infrastructure/platform with real ownership of production systems.
- Strong Python. Not just scripting: you write clean, tested, maintainable pipeline and tooling code that other engineers build on.
- Deep AWS experience in compute, networking, IAM, and storage, and strong opinions about cost.
- A track record of taking rough prototypes (ideally ML/research code) to production
- Hands-on Docker/containerization depth: you debug CUDA base-image conflicts and dependency hell without flinching; ECS/EKS or other orchestration experience.
- Terraform (or equivalent IaC) is used seriously in a team across environments.
- Working GPU knowledge: what saturates a GPU, what leaves it idle, and how instance choice and batching change the bill.
- Cost-optimization instinct: spot strategies, right-sizing, storage tiering, and the discipline to measure before and after.
- Systems thinking and bias to ship: you'd rather run it on real data today and iterate than perfect it in isolation.
Nice to Have:
- Experience with ML labelling/inference pipelines, video, or multimodal sensor data.
- Exposure to computer vision workloads (depth estimation, pose estimation, SLAM/VIO).
- EKS/Kubernetes at scale; Ray or other distributed-compute frameworks.
- CI/CD for container-heavy repos (CodeBuild, GitHub Actions).
Our Stack Languages:
- Python (primary; you must be genuinely strong here), Bash, and Go/Rust a plus.
- Orchestration: AWS Step Functions, AWS Batch, Lambda.
- Containers: Docker, ECR; multi-stage GPU image builds; container orchestration concepts (ECS/EKS) etc.
- IaC: Terraform (modules, multi-environment), GPU/ML runtime: NVIDIA CUDA/cuDNN, PyTorch deployment environments, GPU instance families on AWS, spot vs. on-demand economics.
- Data: S3 at multi-TB scale, structured artifact layouts, dataset versioning, high-throughput transfer.
- Observability: CloudWatch logs/metrics/alarms, cost attribution and reporting.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
653,119 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Free forever. No card. Under a minute.
Your match
How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.
Recommended for you based on this role
Similar stack
Same company
Bengaluru
Apply
Global Specialist Solution Architect - Lightwell
8 hours ago
≈ $94k – $211k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Sydney • Melbourne
Python
Go
JavaScript
Java
Java
Maven
Frontend
npm
DevOps
Ansible
Red Hat
OpenShift
CI/CD
Kubernetes
Platform Engineering
Cybersecurity
SBOM
SLSA
Apply
Global Specialist Solution Architect - Lightwell
8 hours ago
≈ $56k – $126k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Tokyo • Sydney
Python
Go
JavaScript
Java
Java
Maven
Frontend
npm
DevOps
Ansible
Red Hat
OpenShift
CI/CD
Kubernetes
Platform Engineering
Cybersecurity
SBOM
SLSA
Apply
Industrial Engineer
8 hours ago
≈ $10k – $23k per year (Estimated) • In office • Full-Time • 1+ year exp • Bachelor's Degree • Bengaluru
Python
SQL
Analytics
Microsoft Excel
Design
AutoCAD
Apply
Sr Engineer - US
8 hours ago
$98k – $176k per year • In office • Full-Time • 8+ years exp • Brooklyn Park
Python
JavaScript
PHP
Ruby
C#
Node JS
Ruby
Ruby on Rails
C#
.NET
Node JS
Express
Databases
MySQL
PostgreSQL
Oracle
Frontend
Bootstrap
React.js
JQuery
Sass
DevOps
Nginx
Analytics
A/B Testing
Apply
Research Engineer
4 days ago
In office • Master's Degree • Bengaluru
Python
AI/ML
Multimodal AI
Computer Vision
VLM
PyTorch
Vision-Language-Action
DevOps
AWS
Robotics
ROS
LeRobot
Foxglove Studio
SLAM
Visual-Inertial Odometry
Apply
MLOps Engineer
4 days ago
In office • Bengaluru
Python
Rust
Bash
AI/ML
CUDA Toolkit
Multimodal AI
Computer Vision
PyTorch
MediaPipe
Ray
CUDA
cuDNN
DevOps
Terraform
GitHub Actions
CI/CD
AWS
Docker
Kubernetes
Amazon EKS
AWS Lambda
Amazon S3
IAM
Amazon ECS
Amazon CloudWatch
AWS Step Functions
Robotics
Localization
Visual-Inertial Odometry
Apply
Senior Data Engineer
6 days ago
≈ $27k – $55k per year (Estimated) • In office • 4+ years exp • Bengaluru
Python
SQL
Python
FastAPI
Databases
PostgreSQL
Delta Lake
DynamoDB
Apache Hudi
AI/ML
Airflow
Multimodal AI
Ray
Hugging Face
DevOps
SLURM
AWS
Amazon S3
IAM
Analytics
ETL/ELT
AWS Glue
Apply
Infrastructure / Platform Engineer (ML, Data)
20 days ago
In office • Bengaluru
Python
Rust
Bash
AI/ML
CUDA Toolkit
Multimodal AI
Computer Vision
PyTorch
MediaPipe
Ray
CUDA
cuDNN
DevOps
Terraform
GitHub Actions
CI/CD
AWS
Docker
Kubernetes
Amazon EKS
AWS Lambda
GitHub
Amazon S3
IAM
Amazon ECS
Amazon CloudWatch
AWS Step Functions
Robotics
Localization
Visual-Inertial Odometry
Apply
Senior Analyst- Risk Reporting
1 day ago
≈ $20k – $46k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
C#
C++
Databases
Oracle
Analytics
Power BI
Apply
Senior Backend Software Engineer
1 day ago
≈ $27k – $62k per year (Estimated) • Remote/Hybrid • Full-Time • Pune • Bengaluru
Java
SQL
Java
Spring Boot
Databases
PostgreSQL
Redis
DevOps
CI/CD
AWS
AWS Lambda
Amazon S3
Amazon ECS
API Gateway
Management
Agile
Apply
Revenue Manager
1 day ago
In office • 8+ years exp • Bachelor's Degree • Bengaluru
Databases
ElasticSearch
Apply
Workplace Host
1 day ago
≈ $13k – $32k per year (Estimated) • In office • Full-Time • 2+ years exp • Bengaluru
Apply
Procure to Pay Operations New Associate
1 day ago
≈ $15k – $35k per year (Estimated) • In office • Full-Time • Bengaluru
Apply
This is one of many
653,119 more open roles from verified company boards, updated every day.

