721,406open jobs
43,048companies
102,069added this week
Browse all
Salary
$153k – $319k per year (Estimated)
Location
In office (San Francisco, Singapore)
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 24, 2026. First seen by Alion on Sep 16, 2026.

Overview
Company
Impact
Profile match

HUD

About HUD

HUD is building infrastructure to create RL training data and evals for frontier AI agents, as well as a marketplace to sell these to frontier labs through the HUD marketplace. Our platform is used by frontier labs, Fortune 500 companies, and startups. We’ve raised $16M from top VCs and were YC W25.

About the role

We’re looking for a Technical Program Manager to own complex, time-sensitive data and evaluation programs for frontier AI labs, data vendors, and internal teams. You’ll turn ambiguous technical asks into executable programs, coordinate the people and dependencies required to deliver them, and ensure the resulting data is genuinely useful for evaluating and improving LLMs and agents. You’ll also manage vendors and build the processes, tooling, and operating rhythms that allow HUD to deliver reliably at increasing scale.

Responsibilities

  • Own data and evaluation programs end-to-end, from initial scoping and requirements gathering through production, quality assurance, delivery, and retrospective

  • Translate ambiguous requests from AI labs and internal teams into clear specifications such as milestones, owners, dependencies, acceptance criteria, etc.

  • Maintain and improve data quality procedures using both quantitative signals and qualitative inspection, including analyzing task coverage, difficulty, reward distributions, etc.

  • Maintain and improve delivery pipelines at scale, which includes managing external vendors throughout the delivery lifecycle

  • Streamline the logistics behind data production and evaluation programs by improving workflows, documentation, automation, and cross-functional handoffs

  • Partner with research, platform, marketplace teams to translate learnings and feedback into product improvements

Experience

You may be a good fit if you have:

  • Experience owning complex technical programs in AI/ML, data operations, support engineering, applied research, forward-deployed engineering, or a similarly cross-functional environment

  • Strong quantitative and qualitative judgment about data-you can move between aggregate metrics and individual examples to determine whether a dataset or evaluation is useful, reliable, and representative

  • Experience maintaining and improving complex processes, especially within project delivery or data contexts

  • Demonstrated ability to bring structure to ambiguous problems, maintain momentum as requirements change, and make sound tradeoffs without waiting for a perfect specification

  • Strong written and verbal communication skills, including the ability to turn messy technical information into clear decisions, owners, and next steps

Strong candidates may also:

  • Experience with RLHF, reinforcement-learning environments, agent evaluations, human-in-the-loop data pipelines, annotation, or expert data collection

  • Experience supporting frontier AI labs, technical enterprise customers, or research teams with urgent and evolving requirements

  • Early-stage startup experience with ability to work independently in fast-paced environments

We prioritize technical aptitude and learning potential over years of experience. Motivated candidates are encouraged to apply even if they don't meet all criteria.

Team & company details

  • Team Size: ~25 people currently, mostly full-time in-person, but some remote.

  • Our team: Our team includes 4 International Olympiad medalists (IOI, ILO, IPhO), serial AI startup founders, and researchers with publications at ICLR, NeurIPS, etc.

  • Company stage: We have 8 figures in funding and are scaling profitably and quickly to meet very strong demand.

Logistics

  • Employment: Full-time.

  • Location: We have offices in San Francisco or Singapore but are open to remote candidates who can work hours that 70-80% overlap with either San Francisco or Singapore time zones.

  • Visa Sponsorship: We provide support for relocation and visas for strong full-time candidates to the US or Singapore.

  • Timeline: Applications are rolling. The process is 2 technical interviews and a 2-3 day work trial.

What we offer

  • Competitive compensation

  • 100% covered top-of-the-line medical, dental, and vision from Blue Shield of CA (US employees)

  • Lunch and dinner when you’re in the office (in-office employees)

  • Company-wide holiday break (Christmas Eve to New Year’s Day) on top of PTO and paid holidays

  • Other perks including an Equinox membership, 401k, and commuter benefits (US employees)

  • Unlimited* access to tokens for ChatGPT, Claude Code, Cursor, etc. *By unlimited, we mean no one on our token usage leaderboard has ever hit a limit. So we have no idea what the limit is.

Due to high volume, we may not actively respond to every application, but feel free to contact us at [email protected] or elsewhere if we missed your application!

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
721,406 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
Data Scientist 2 days ago
$43k – $116k per year (Estimated) • Remote • Full-Time • Bachelor's Degree
Python
AI/ML
Copilot
Cursor
Weights & Biases
Claude Code
LoRA
MLFlow
Vertex AI
Fine-tuning
Scikit-learn
Multimodal AI
PEFT
QLoRA
Kubeflow
Transformers
NumPy
PyTorch
LLM
OpenAI
Anthropic
Amazon SageMaker
OpenAI Codex
Machine Learning
DevOps
GCP
Azure
Git
AWS
Cybersecurity
SOC 2
HIPAA
Apply
Data Engineer 2 days ago
$42k – $115k per year (Estimated) • Remote • Full-Time • Bachelor's Degree
Python
SQL
Databases
PostgreSQL
Snowflake
Databricks
pgvector
Pinecone
FAISS
Qdrant
Apache Kafka
Google BigQuery
Amazon Redshift
Azure SQL Database
BigQuery
AI/ML
Copilot
Cursor
Spark
Claude Code
Dagster
dbt
Prefect
Flink
RAG
OpenAI
Anthropic
OpenAI Codex
DevOps
Terraform
GitHub Actions
Azure
CI/CD
Git
AWS
Docker
Bicep
Cybersecurity
SOC 2
HIPAA
Analytics
Dimensional Modeling
Apply
$52k – $136k per year (Estimated) • Equity • In office • Full-Time • 10+ years exp • Bachelor's Degree • Hsinchu
AI/ML
Copilot
ChatGPT
Apply
$77k – $102k per year • In office • Bachelor's Degree • New York
AI/ML
Copilot
ChatGPT
DevOps
Windows
TCP/IP
DNS
DHCP
VPN
Cybersecurity
Active Directory
Management
ServiceNow
Microsoft Teams
Service Desk
Microsoft Office
Apply
$77k – $159k per year (Estimated) • Equity • In office • Full-Time • 18+ years exp • Bachelor's Degree • Hsinchu
AI/ML
Copilot
ChatGPT
Cybersecurity
CAPA
Management
Agile
Apply
$162k – $349k per year (Estimated) • In office • Full-Time • San Francisco
Python
AI/ML
Cursor
ChatGPT
Claude Code
AI Agents
LLM
Synthetic Data
Apply
$159k – $293k per year (Estimated) • In office • Full-Time • San Francisco
AI/ML
Cursor
ChatGPT
Claude Code
AI Agents
DevOps
Terraform
Helm
CI/CD
AWS
Docker
Kubernetes
Amazon EKS
Amazon EC2
Amazon S3
IAM
Apply
Open Role 7 days ago
In office • Full-Time • San Francisco
AI/ML
Cursor
ChatGPT
Claude Code
AI Agents
Apply
$139k – $277k per year (Estimated) • In office • Full-Time • San Francisco
AI/ML
Cursor
ChatGPT
Claude Code
AI Agents
DevOps
AWS
GitHub
Cybersecurity
SOC 2
Threat Modeling
Apply
$112k – $298k per year (Estimated) • In office • Full-Time • San Francisco
AI/ML
Cursor
ChatGPT
Claude Code
AI Agents
Apply
$46k per year • In office • Full-Time • San Francisco
Management
Agile
Apply
$197k – $314k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • San Francisco • Washington
AI/ML
Cursor
Claude Code
Function Calling
AI Agents
Langfuse
LangSmith
LLM
Braintrust
OpenAI Codex
Agentforce
Replit
LLM Evaluation
Tool Use
Management
Slack
Apply
$229k – $345k per year • Equity • Remote/Hybrid • Full-Time • 7+ years exp • San Francisco • Milpitas • Mountain View • Los Altos • Fremont
DevOps
BGP
OSPF
Apply
$139k – $235k per year • In office • Full-Time • San Francisco
Apply
$139k – $299k per year (Estimated) • Remote • Full-Time • San Francisco
AI/ML
Cursor
Claude Code
Time Series Forecasting
Physical AI
Machine Learning
Apply
See all jobs
This is one of many
721,406 more open roles from verified company boards, updated every day.