665,767open jobs
38,907companies
100,291added this week
Browse all
Salary
$156k – $342k per year (Estimated)
Location
In office (New York)
Employment
Full-Time
Overview
Company
Impact
Profile match
Mecka is the data and deployment layer for physical AI — capturing, structuring, and evaluating real-world activity into datasets robotics and AI teams can train on.

About Mecka AI

Mecka AI is building the data infrastructure layer for robotics and embodied AI.

We design and operate global systems for data capture, data labeling, and hardware-enabled workflows used by leading AI labs and robotics companies to train and validate humanoid and embodied AI systems.

We work closely with frontier robotics teams to bridge real-world data, simulation, learning-based systems, and deployed hardware.

About Mecka AI

Mecka AI is building the data infrastructure layer for robotics and embodied AI. We partner with leading AI labs and robotics companies to deliver high-quality, real-world datasets used to train, evaluate, and deploy robotic systems-where model performance is dictated by data quality.

The Role

While our existing perception division handles state estimation and spatial mapping, this role is dedicated to one of the most critical bottlenecks in embodied AI: dexterous manipulation and human-object interaction. We are hiring a Research Scientist to architect and train proprietary foundation models from scratch focused on 3D hand tracking and articulated pose estimation.

Your core mandate is twofold: building our in-house equivalents to cutting-edge 3D hand and mesh recovery architectures, and developing highly robust interaction models tailored for the heavily occluded, chaotic domain of egocentric manipulation. Beyond these core pillars, you will serve as a lead problem-solver for emergent perception challenges as our hardware and downstream robotics needs evolve.

To achieve this, we can provide a massive, continuous stream of high-quality, proprietary ground-truth manipulation data captured by our infrastructure. You will use this data advantage to train networks that surpass current public baselines, owning the complete hand-object perception loop for our data engine.

What You'll Work On

Architecting Proprietary Articulation Models

  • Zero-to-One Model Development: Design, implement, and train state-of-the-art networks for 3D hand pose estimation, dense mesh recovery, and kinematic tracking.

  • Large-Scale Distributed Training: Scale multi-view and temporal ML architectures across multi-GPU clusters to handle massive, multi-modal datasets of human hands in action.

  • Loss & Architecture Innovation: Push the boundaries of current paradigms by developing novel loss functions that enforce biomechanical constraints, temporal smoothness, and physical plausibility.

Egocentric Hand-Object Interaction (HOI)

  • Egocentric Manipulation Modeling: Build and train custom architectures capable of handling the extreme motion blur, severe self-occlusion, and rapid rotations inherent in first-person object manipulation.

  • Dynamic Scene Understanding: Use your models to track objects through complex grasps, segment tools from hands, and map contact points and forces to provide rich regularization for downstream action-conditioned robotics models.

Emergent Perception R&D

  • Rapid Prototyping: Tackle novel, unmapped AI challenges as they arise. You will rapidly prototype and deploy new models for tasks spanning tactile-visual fusion, fine-grained action segmentation, and novel hardware sensor integrations.

  • Agile Problem Solving: Pivot to resolve sudden algorithmic bottlenecks in the data engine, adapting the latest research to unblock new product capabilities for our robotics customers.

Dense Contact & Physics-Aware Tracking

  • Interaction Integration: Connect the outputs of your foundational tracking models into highly optimized pipelines that reason about physical contact surfaces and object affordances, directly bridging the gap between human video data and robotic control policies.

Who You Are

Required Background

  • Deep expertise in Deep Learning, 3D Computer Vision, and specifically Articulated Tracking / Hand Pose Estimation.

  • Proven experience training large-scale vision models from scratch, not just running inference or fine-tuning existing checkpoints.

  • Strong theoretical and practical understanding of parametric hand models, inverse kinematics, and dense mesh estimation.

  • Mastery of PyTorch and deep learning scaling frameworks.

  • Experience handling and curating massive, multi-terabyte image and video datasets for training.

  • Comfortable operating in a fast-paced environment where priorities can shift rapidly to capitalize on new research or hardware capabilities.

Warning: Research Scientist positions require hyper-specific expertise. Please limit your applications to one research role. Applying to multiple Research Scientist positions suggests a lack of focus and may result in the rejection of all submissions. You may, however, apply to other non-research roles alongside your research application.

Strong Signals:

  • First-author publications in top-tier venues (CVPR, ICCV, ECCV, NeurIPS) focusing on 3D hand tracking, hand-object interaction (HOI), dexterous manipulation, or human mesh recovery.

  • Specific experience working with massive egocentric manipulation datasets (e.g., Ego4D, Ego-Exo4D, DexYCB, Epic-Kitchens) and solving the unique optimization challenges they present.

  • Experience writing custom CUDA kernels to accelerate 3D operations, differentiable rendering of meshes, or collision/contact computation.

Why This Role?

  • The Data Advantage: You will have access to a scale and quality of proprietary spatial and temporal ground truth for human manipulation that most academic researchers only dream of.

  • Pure R&D & Model Ownership: You are not maintaining legacy systems; you are given a blank slate and the compute resources to build the state-of-the-art.

  • High Impact: The kinematic priors and interaction models you architect will directly define how the next generation of embodied AI agents learn to physically grasp and manipulate the world.

Inclusive Hiring at Mecka

We are committed to creating an inclusive and supportive candidate experience. Should you require any accommodation whatsoever during the interview process, please inform us without any hesitation. Mecka is dedicated to ensuring equal treatment and opportunity in all phases of recruitment, selection, and employment, in compliance with employment law. We do not discriminate based on gender, race, religion, national origin, ethnicity, disability, gender identity/expression, sexual orientation, veteran or military status, or any other protected category. Mecka is proud to be an equal opportunity employer, fostering a culture of inclusivity and maintaining a work environment that is free from discrimination, harassment, and retaliation.

Use of Artificial Intelligence in Recruitment

Mecka uses artificial intelligence (AI) responsibly to support administrative and efficiency-focused aspects of our recruitment process. This includes activities such as drafting job descriptions, generating interview questions, note-taking and recordings, and supporting sourcing and scheduling workflows. All candidate evaluations, interviews, and hiring decisions are made by members of the Mecka team. While AI tools may assist with screening and assessment, they do not replace human judgment in selection decisions. Our use of AI is intended to streamline routine tasks, improve consistency, and enhance the overall candidate experience. We are committed to upholding principles of fairness, transparency, and accountability in all hiring activities. Mecka regularly reviews its recruitment practices to mitigate bias and to ensure alignment with applicable laws and evolving best practices.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
665,767 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
New York
$97k – $122k per year • Equity • Remote/Hybrid • Full-Time • Bachelor's Degree • Toronto
Python
AI/ML
llama.cpp
LoRA
vLLM
Fine-tuning
Prompt Engineering
Function Calling
Langfuse
LangSmith
PEFT
TGI
LLM
RAG
DPO
SFT
Tool Use
DevOps
CI/CD
Git
Docker
GitHub
Apply
In office • TS/SCI • 8+ years exp • Bachelor's Degree
Python
Java
AI/ML
LangChain
vLLM
AI Agents
LiteLLM
RAG
Edge AI
DevOps
OpenTelemetry
Prometheus
CI/CD
AWS
Kubernetes
Grafana
Vector
Apply
In office • TS/SCI • 12+ years exp • Bachelor's Degree
Python
Java
TypeScript
AI/ML
AI Agents
LLM
DevOps
CI/CD
Apply
Software Engineer 7 hours ago
In office • TS/SCI • 3+ years exp • Bachelor's Degree
Python
JavaScript
AI/ML
Claude Code
AI Agents
Frontend
Vue.js
React.js
DevOps
CI/CD
Git
Docker
Kubernetes
Management
Confluence
Jira
Apply
In office • TS/SCI • 8+ years exp • Bachelor's Degree
Python
AI/ML
LangChain
vLLM
AI Agents
LiteLLM
RAG
DevOps
OpenTelemetry
Prometheus
CI/CD
AWS
Kubernetes
Grafana
Vector
Apply
$120k – $150k per year • In office • Full-Time • 5+ years exp • New York
AI/ML
Embodied AI
Apply
Growth Lead 5 days ago
$150k – $180k per year • In office • Full-Time • New York
AI/ML
Embodied AI
Apply
$115k – $144k per year • In office • Full-Time • Shenzhen • Toronto
AI/ML
OpenCV
Embodied AI
Robotics
Sensor Fusion
Apply
$93k – $129k per year • Remote/Hybrid • Full-Time • Toronto
Python
JavaScript
TypeScript
C
Python
FastAPI
C
FFmpeg
AI/ML
Jupyter Notebook
Fine-tuning
Computer Vision
ONNX
TensorRT
PyTorch
Human-in-the-Loop
Embodied AI
Frontend
Vue.js
React.js
DevOps
gRPC
GitHub
Web3
Consensus
Apply
$150k – $185k per year • Remote/Hybrid • Full-Time • New York
Python
C
C++
C
FFmpeg
C++
PyTorch C++
AI/ML
OpenCV
Multimodal AI
Computer Vision
Gradio
PyTorch
Embodied AI
Robotics
Ceres Solver
GTSAM
SLAM
Sensor Fusion
Visual-Inertial Odometry
COLMAP
Apply
$161k – $296k per year (Estimated) • In office • Full-Time • 10+ years exp • Bachelor's Degree • New York • Madison
Python
Java
Apply
$60k per year • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • New York
Management
Google Workspace
Apply
$172k – $346k per year (Estimated) • In office • Full-Time • New York
AI/ML
AI Agents
Management
Notion
Apply
$195k – $370k per year (Estimated) • In office • Full-Time • New York • Philadelphia
Apply
$96k – $153k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Milpitas • Charlotte • Raleigh • Tampa • Baltimore
Apply
See all jobs
This is one of many
665,767 more open roles from verified company boards, updated every day.