368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$150k – $250k per year
Location
In office (San Francisco)
Employment
Full-Time
Overview
Company
Impact
Profile match
Visual understanding for autonomy.. Manage fleet-scale visual robotics data. Mine for critical long-tail scenarios and curate datasets. Dataloading into training. Monitor live deployments.

About Eventual

Every breakthrough Physical AI system - humanoid robots, autonomous vehicles, video generation models - is trained on petabytes of video, lidar, radar, and sensor data. But today's data platforms (Databricks, Snowflake) were built for spreadsheet-like analytics, not the multimodal corpora that power AI. As a result, robotics and video-AI teams iterate on model improvement about once a week. Most of that week isn't training - it's finding the right data: writing CV heuristics over raw footage, paying annotators for edge cases, hand-curating clips before a cluster ever spins up. GPU bandwidth has grown 2-3× per generation. Storage and pipelines haven't. The gap widens every year.

Eventual was founded in 2022 to close it. Our open-source engine, Daft, is the distributed data engine purpose-built for multimodal AI - already running 2 PB/day at Amazon, 60-100 PB at another FAANG company, and in production at Mobileye, TogetherAI, and CloudKitchens. We are building a video-native index on top of our engine for Physical AI that collapses the data iteration loop. Describe the dataset you want, get a curated table in minutes, feed it to your GPUs at line rate. One iteration per day becomes the norm.

We're building this in partnership with the top PhysicalAI labs and public AI infrastructure companies today. We have raised $30M from Felicis, CRV, Microsoft M12, Citi, Essence, Y Combinator, Caffeinated Capital, Array.vc, and angels from the co-founders of Databricks and Perplexity. We've assembled a world-class team from AWS, Render, Pinecone and Tesla. We have spent our careers powering the last generation of PhysicalAI in self-driving, and are excited to now do this for the next.

Join our small (but powerful!) team working together 4 days/week in our SF Mission district office.

Your Role

As a Research Engineer on the Visual Understanding team, you'll own the layer that makes petabytes of video queryable by content. Physical AI teams have raw footage, lidar, radar, and sim outputs scattered across object stores with no way to find what they need without weeks of human annotation. We change that economics: we run vision-language models over every clip in a corpus along axes the customer cares about (gripper type, failure mode, object class, scene, motion density), so a researcher can ask "left-arm grasp failures on deformable objects" and get a curated dataset in minutes.

You'll define the roadmap for our visual understanding capabilities, train and select the models that make corpus-scale annotation tractable at single-digit cents per hour of video, and build the rich datasets that go on to train customer models. This is a research engineering role - meaning you'll read papers and run experiments, but you ship to production and your work is judged by what it does for customer training runs.

Key Responsibilities

  • Own the visual understanding roadmap end-to-end: from picking the model family for a customer's taxonomy to landing it in production inference at corpus scale.

  • Train, fine-tune, and evaluate VLMs, VQA models, embedding models, and convolutional perception models against customer datasets and benchmarks.

  • Drive down per-clip annotation cost - model selection, distillation, batching, decode pipelining - so "annotate every clip in a 10K-hour corpus" stays economical.

  • Build the rich, queryable datasets that customers train on: design taxonomies with researchers, instrument quality, version the outputs.

  • Partner with the dataloading and storage teams so visual understanding outputs flow into the index and on to the GPU without re-engineering.

  • Work directly with researchers at our partner labs - your shortest feedback loop is their next training iteration.

What we look for

  • Strong familiarity with modern vision and multimodal models - convolution nets, VLMs, VQA, embeddings - and a sense for the SOTA that's actually deployable today vs. on a leaderboard.

  • Experience running these models at scale on real video and sensor data, ideally for perception tasks (detection, tracking, segmentation, retrieval, captioning).

  • Background from a perception team at a self-driving, robotics, or visual-data company - or equivalent depth from a research lab.

  • Comfortable with cloud infrastructure and large-scale data processing - you don't need to be a distributed-systems engineer, but you've shipped jobs that ran on thousands of GPU-hours of video.

  • Bias toward data and infrastructure: you reach for "annotate the whole corpus" before "fine-tune another model."

Nice to have

  • Experience training vision or multimodal models from scratch (not just calling APIs).

  • ML/AI research background - papers, citations, or a research org on your resume.

  • Hands-on time with big-data frameworks like Spark, Ray, or Daft.

  • Worked on embeddings, retrieval, or content-aware search at scale.

  • Experience designing labeling taxonomies or running annotation programs.

Perks & Benefits

  • In-person, tight-knit team - 4 days/week in our SF Mission office.

  • Competitive comp and meaningful startup equity.

  • Catered lunches and dinners for SF employees.

  • Commuter benefit.

  • Team-building events and poker nights.

  • Health, vision, and dental coverage.

  • Flexible PTO.

  • Latest Apple equipment.

  • 401(k) plan with match.

If you're excited about being on the team that turns petabytes of raw video into the training data for the next generation of Physical AI, we'd love to talk.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$98k – $195k per year (Estimated) • In office • Full-Time • 7+ years exp • Wellington
Java
Python
SQL
Java
Spring Boot
Databases
Apache Kafka
Databricks
Neo4j
AI/ML
Flink
Spark
Frontend
GraphQL
DevOps
Azure
CI/CD
Datadog
Dynatrace
Kibana
Kubernetes
OpenShift
Platform Engineering
Splunk
Amazon ECS
Apply
$84k – $178k per year (Estimated) • In office • Full-Time • 10+ years exp • Wellington
Java
Python
DevOps
Ansible
AWS
Azure
CI/CD
Docker
GCP
Helm
Kubernetes
Platform Engineering
Prometheus
Service Mesh
Terraform
GitLab
IAM
Apply
Platform Engineer 1 day ago
$87k – $140k per year • In office • Full-Time • 3+ years exp • Berlin
Databases
PostgreSQL
Redis
DevOps
AWS
Azure
Bicep
CI/CD
Docker
GCP
GitHub Actions
Kubernetes
OpenShift
Terraform
GitHub
Apply
Founding Engineer 1 day ago
$81k – $116k per year • In office • Full-Time • Bachelor's Degree • Munich
JavaScript
Python
TypeScript
Databases
MySQL
PostgreSQL
Frontend
Next.js
React.js
Tailwind CSS
DevOps
AWS
Azure
CI/CD
Docker
GCP
Grafana
Kubernetes
OpenTelemetry
Prometheus
Apply
Product Engineer 1 day ago
$105k – $128k per year • In office • Full-Time • Leipzig
TypeScript
Databases
PostgreSQL
AI/ML
AI Agents
DevOps
AWS
Apply
$150k – $250k per year • In office • Full-Time • 3+ years exp • San Francisco
C++
Rust
Databases
Apache Iceberg
Databricks
Delta Lake
Google BigQuery
Pinecone
PostgreSQL
Snowflake
AI/ML
Hadoop
Multimodal AI
Perplexity
Ray
Spark
Together AI
DevOps
AWS
Amazon S3
Apply
$149k – $303k per year (Estimated) • In office • Full-Time • San Francisco
C++
Go
Python
Rust
Databases
Databricks
Pinecone
Snowflake
AI/ML
Computer Vision
Copilot
Multimodal AI
Perplexity
Together AI
DevOps
AWS
Kubernetes
SLURM
GitHub
Apply
$180k – $210k per year • Equity • In office • Full-Time • San Francisco
Node JS
JavaScript
Databases
PostgreSQL
DevOps
PagerDuty
Web3
TRM Labs
Management
Slack
Apply
$252k – $335k per year • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco
AI/ML
ChatGPT
Human-in-the-Loop
OpenAI
OpenAI Codex
DevOps
SLI/SLO/SLA
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
$160k – $283k per year • Equity • In office • 5+ years exp • San Francisco
AI/ML
AI Agents
Apply
$185k – $385k per year • Remote/Hybrid • Full-Time • 5+ years exp • San Francisco
JavaScript
Python
Databases
MySQL
PostgreSQL
AI/ML
OpenAI
Frontend
React.js
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.