368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$180k – $450k per year
Location
In office (San Jose)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match
Hark was an online digital entertainment platform best known for its extensive library of short audio soundbites, video clips, and pop culture quotes. Launched in 2007, the website allowed users to browse, create, and share playable soundboards featuring memorable lines from movies, television shows, and political figures. While it grew into a popular destination for viral sound clips during the late 2000s and early 2010s, the platform has since ceased its original operations.

About Hark

Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persistent memory.

We're pairing that intelligence with next-generation hardware to create a universal interface between humans and machines. While today's AI largely operates through chat boxes and decade-old devices, Hark is focused on what comes next: agentic systems that interact naturally with people and the real world.

To get there, we're developing multimodal models and next-generation AI hardware together - designed from the ground up as a single, unified interface for a new era of intelligent systems.

About the Role

We are looking for a Member of Technical Staff - Mid-Training to lead the development of training strategies that bridge pre-training and post-training, shaping how models acquire reasoning, planning, and tool-use capabilities at scale.

This role sits at the core of model capability development-defining how data, algorithms, and systems interact to unlock the next frontier of agent behavior.

Responsibilities

  • Design and implement mid-training strategies to improve agent capabilities such as reasoning, planning, tool use, and long-horizon decision-making.
  • Scale synthetic data generation pipelines (e.g., coding, agent trajectories, multimodal data) and optimize data mixtures to improve downstream RL performance.
  • Build and optimize distributed training pipelines for large models, ensuring efficiency, stability, and scalability across GPU clusters.
  • Develop and iterate on evaluation frameworks to measure model capability (e.g., task success, reasoning quality, tool use accuracy) and guide training improvements.
  • Conduct rigorous experimentation and ablations to understand training dynamics, scaling behavior, and bottlenecks.
  • Collaborate cross-functionally with pre-training, post-training, and product teams to align model development with real-world agent use cases.
  • Drive technical innovation in areas such as long-context learning, data distillation, and training efficiency, while contributing to the overall model roadmap.

Requirements

  • Strong background in machine learning, with hands-on experience training or fine-tuning large models - LLMs, multimodal, or equivalent systems.
  • Deep understanding of reinforcement learning: policy optimization, reward design, exploration, and the interplay between environment design and agent behavior.
  • Experience building or working within simulation or execution environments (e.g., code interpreters, sandboxed execution, game environments, robotics simulators).
  • Proven ability to design and execute rigorous experiments, with strong intuition for diagnosing training failures and scaling bottlenecks.
  • Proficiency in Python and PyTorch; comfort working across research and systems code.
  • Ability to work in a fast-moving, research-forward environment where the right approach is often unknown at the outset.

We expect strong candidates to come from a range of backgrounds - RL research, robotics, competitive programming systems, compilers, formal methods, or large-scale ML - rather than post-training specifically. The field is new enough that directly relevant experience is rare; what matters is depth, rigor, and transferability.

Bonus Qualifications

  • Experience with mid-training, post-training, or agent-focused model development or coding LLM training.
  • Familiarity with synthetic data generation, trajectory-based training, or coding/model distillation pipelines.
  • Experience training or scaling large models (100B+ parameters or equivalent systems).
  • Background in reinforcement learning, decision-making systems, or agent frameworks.
  • Contributions to open-source ML systems or publications at top conferences (ICML, NeurIPS, ICLR, ACL, etc.).
  • Experience optimizing distributed training systems (GPU utilization, memory efficiency, communication).

Compensation

The US base salary range for this full-time position is between $180,000 - $450,000 annually.

The pay offered for this position may vary based on several individual factors, including job-related knowledge, skills, and experience. The total compensation package may also include additional components and benefits depending on the specific role. This information will be shared if an employment offer is extended.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Jose
Data Engineer 3 10 hours ago
$21k – $52k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Bengaluru
Java
Python
Scala
SQL
Java
Maven
Python
pySpark
Databases
Apache Kafka
Databricks
Snowflake
AI/ML
Hadoop
Spark
DevOps
Azure
Analytics
Tableau
ETL/ELT
Apply
Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Istanbul
JavaScript
TypeScript
AI/ML
Fine-tuning
Frontend
Angular
RxJS
Sass
DevOps
Azure
Bitbucket
Git
Apply
$22k – $48k per year (Estimated) • Remote • Contractor • Chelyabinsk
Python
SQL
Apply
$16k – $40k per year (Estimated) • Remote/Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • Gurgaon
Python
SQL
Analytics
Power BI
Tableau
ETL/ELT
Apply
$14k – $32k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Apply
$180k – $450k per year • In office • Full-Time • San Jose
Python
AI/ML
Fine-tuning
Knowledge Distillation
LLM
Multimodal AI
PyTorch
Reinforcement Learning
RLHF
Synthetic Data
DPO
GRPO
Post-training
PPO
AI Agents
Function Calling
Robotics
Imitation Learning
Reinforcement Learning
Apply
Interface Designer 5 days ago
$150k – $300k per year • In office • Full-Time • 5+ years exp • San Jose
AI/ML
Multimodal AI
AI Agents
Apply
$180k – $450k per year • In office • Full-Time • San Jose
AI/ML
DeepSpeed
LLM
Multimodal AI
Synthetic Data
Megatron-LM
AI Agents
Apply
$180k – $450k per year • In office • Full-Time • San Jose
AI/ML
DeepSpeed
Fine-tuning
Multimodal AI
Reinforcement Learning
RLHF
DPO
FSDP
GRPO
Megatron-LM
Post-training
PPO
SFT
AI Agents
Apply
$180k – $450k per year • In office • Full-Time • San Jose
AI/ML
Multimodal AI
Speech Recognition
Synthetic Data
Text-to-Speech
AI Agents
Apply
$147k – $265k per year (Estimated) • In office • Full-Time • Folsom • San Jose
Apply
$117k – $255k per year (Estimated) • In office • Full-Time • 8+ years exp • Richardson • Boise • Folsom • San Jose
Verilog
AI/ML
Claude
Apply
$124k – $208k per year • In office • Full-Time • 2+ years exp • Bachelor's Degree • Austin • San Jose
C++
Python
SystemVerilog
Chips/EDA
Formal Verification
Apply
$116k – $253k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Boise • San Jose
Apply
$68k – $85k per year • In office • Full-Time • Master's Degree • San Jose
Python
AI/ML
AI Agents
DevOps
Amazon EC2
AWS
AWS Lambda
Bitbucket
CI/CD
CloudFormation
Docker
Git
Kubernetes
Terraform
Amazon S3
IAM
HPC
Cybersecurity
Least Privilege
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.