373,573open jobs
9,691companies
51,475added this week
Browse all
Salary
$179k – $334k per year (Estimated)
Location
Remote (United States)
Seniority
Staff · 8+ years exp
Overview
Company
Impact
Profile match
Accelerate development & testing with Tonic.ai. Generate realistic, production-like test data that preserves privacy & compliance in complex environments. Learn more!

About Tonic

Tonic builds the data infrastructure behind modern AI. We generate the synthetic environments that agents are trained and tested in, and we de-identify real enterprise data so it can be used safely in training and evaluation. Eight years in, we work with frontier AI labs pushing the edge of what models can do, and with hundreds of enterprises including Fidelity, Comcast, eBay, and Vanguard, on the data problems that sit at the center of where AI is going next.

About The Role

The models you build here are load-bearing. The environments you generate decide whether an agent is ready to ship or only looked good in a demo. The synthesis and de-identification models you train decide whether a bank can safely put its data near a model at all. And the work spans real range: in one week you might build evaluation that separates the best models from the rest on real tasks, train a synthesis model where both fidelity and downstream utility have to hold, and improve entity detection on messy production data. Real enterprise data, real stakes, and problems that don't have textbook answers yet.

What You'll Do

  • Design and build the systems that generate longitudinally coherent synthetic environments for agent training and evaluation, including persona modeling, task generators, and verifiable ground truth.

  • Build and maintain synthesis models that generate realistic replacement values at very large scale, preserving format, statistical distribution, and semantic consistency so de-identified data stays useful downstream.

  • Train and improve the NER models behind our entity detection, driving accuracy and recall across free text, structured fields, and mixed enterprise data at scale.

  • Build evaluation infrastructure that grades agent outcomes, not just traces, and produces real discrimination between frontier models on real tasks.

  • Fine-tune and evaluate open-weight models on Tonic-generated data, and turn benchmark results into product and research direction.

  • Expand coverage into new domains, languages, and entity types, and handle the long tail of formats and edge cases that real customer data throws off.

  • Own model evaluation across the board: precision and recall on detection, utility preservation on synthesis, and outcome-level grading for agents.

  • Optimize inference so models run efficiently on large volumes of sensitive data inside customer environments.

  • Partner directly with frontier labs and enterprise ML team to turn hard data problems into shipped model improvements.

  • Set technical direction for a small, senior team and raise the bar on rigor, reproducibility, and shipping.

What You'll Bring

  • 8+ years (or PhD with 3+ years) building production ML systems, with real depth in some combination of LLMs, agents, RL, NER, or information extraction.

  • Hands-on experience training and shipping models to production, and a pragmatic bar for quality: you know how to measure it, where it breaks, and when it's good enough to ship.

  • Experience with generative or synthesis models where output fidelity and downstream utility both matter, not just plausibility.

  • Strong software engineering fundamentals. You write code others build on.

  • Fluency with modern training and eval stacks (PyTorch, distributed training, standard agent and benchmark frameworks).

  • Comfort working with messy, sensitive, real-world data and the privacy constraints that come with it.

  • A track record of framing ambiguous problems and driving them to measurable, shipped results.

  • Bonus: synthetic data generation, data privacy or de-identification, or benchmark construction.

Benefits We Offer

  • Competitive salary and equity

  • Unlimited paid time off

  • 401k plan with employer contribution

  • Medical, dental, and vision insurance

  • Generous parental leave policy

  • Remote-friendly work environment

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
373,573 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
In office • Full-Time • Master's Degree • Singapore
C++
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
Agentforce
AI Agents
Edge AI
JAX
Multimodal AI
NLP
Prompt Engineering
PyTorch
TensorFlow
Time Series Forecasting
Marketing
Salesforce
Apply
$35k – $84k per year (Estimated) • In office • Full-Time • 10+ years exp • Bengaluru • Gurgaon
Python
Rust
SQL
Python
FastAPI
AI/ML
Fine-tuning
Hugging Face
Knowledge Distillation
LangChain
LLM
NumPy
Pandas
Prompt Engineering
PyTorch
Quantization
RAG
Recommender Systems
Scikit-learn
Semantic Search
Semantic Search
Sentiment Analysis
TensorFlow
Transformers
vLLM
DevOps
AWS
Azure
CI/CD
GCP
Analytics
A/B Testing
Apply
AI / ML Engineer 9 hours ago
$26k – $72k per year (Estimated) • In office • Full-Time • 3+ years exp • Pune
Python
AI/ML
PyTorch
TensorFlow
Apply
GenAI Developer 1 day ago
$27k – $69k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Master's Degree • Gurgaon
Node JS
Python
JavaScript
AI/ML
AI Agents
Embeddings
Fine-tuning
Hugging Face
LoRA
PEFT
Prompt Engineering
PyTorch
Quantization
RAG
TensorFlow
Transformers
DevOps
Vector
Apply
$30k – $88k per year (Estimated) • In office • Full-Time • Moscow
Bash
Python
Python
Mypy
AI/ML
Claude
Claude Code
Cursor
GigaChat
Hugging Face
LangGraph
LLM
NumPy
OpenAI Codex
PyTorch
RAG
SGLang
vLLM
VLM
LangChain
DevOps
Amazon S3
CI/CD
Docker
Git
Apply
See all jobs
This is one of many
373,573 more open roles from verified company boards, updated every day.