411,234open jobs
14,368companies
73,679added this week
Browse all
Salary
$96k – $144k per year
Location
In office (San Francisco)
Seniority
Intern
Employment
Internship
Overview
Company
Impact
Profile match
Headquartered in San Francisco, California, CTGT is an applied AI research laboratory and enterprise software startup specializing in mechanistic interpretability, representation engineering, and AI governance. Backed by Y Combinator and Google's Gradient Ventures, the firm develops deterministic policy engines and alignment systems that intervene directly at an LLM's neural representation layer during inference.

About CTGT & The Mission

Despite massive investment in commercial AI, organizations often find that demonstrated value is elusive, primarily due to the non-deterministic risk inherent to generative models. CTGT is the deterministic governance layer that enables the most important global institutions to deploy AI workflows with confidence.

Born out of Stanford University research, we provide the control plane that makes it possible. A lightweight, model-agnostic system that enforces policy, prevents drift, and produces auditable decisions in real time. When benchmarked on HaluEval, the CTGT Policy Engine (paired with GPT-120B OSS) outperformed frontier models (Gemini 3 Pro Preview, Claude 4.5 Opus and 4.5 Sonnet) at drastically lower compute cost.

While we sit on the edge of AI research, CTGT brings frontier intelligence into real-world environments. We apply cutting-edge theory directly in production to make large language models more reliable, controllable, and performant in practice.

Our mission is to bring models to the level of performance and accountability required by the Fortune 500. By bridging the gap between LLM capabilities and domain-specific requirements, we unlock the true potential of generative AI to solve the most pressing problems in our world today.

The Role

Not your average fixed-point internship.

Frontier models are now usually right and occasionally confidently wrong, and they cannot tell you which is which. A model that is 95% reliable is useless in the settings we serve, the same way a self-driving car that avoids most accidents is useless. CTGT's research function exists to close that gap. Our founding research stems from feature learning in neural networks, and we use that machinery to extract and steer features at runtime, on open and closed-weight models, without training a new artifact for every behavior.

As a research intern, you will own one hard problem inside this program from end to end. You will not be handed a labeling task or a notebook to babysit. You will take a real research question, like a better way to find what a model represents, intervene on it, or bound how wrong the system can be, design an approach, implement it against real models, and prove or disprove it with evidence that holds up. You will sit directly with the engineers building the Policy Engine, present in our weekly research review, and be expected to form opinions, ask hard questions, and take problems further than they were handed to you.

We hold interns to the standard of a calculation that has to be right, not a demo that usually works; in practice this means limited ground truth, unverifiable intermediate steps, and failure modes that hide in the tails.

What You Will Do

  • Implement and stress-test methods for feature extraction and runtime intervention, from control vectors to activation probes, and make them work repeatably across model families
  • Design evaluations that bound error rather than average it: calibration under imbalanced data, reasoning-trace grading, behavior in verifiable and non-verifiable task regimes
  • Read the relevant literature, decide what actually matters, reproduce it, and push past it
  • Work with engineering to turn a finding into a Policy Engine capability that ships into audited, high-stakes environments
  • Present your progress every week and defend your reasoning

Who You Are

  • Pursuing a degree (Bachelor's through PhD) in computer science, mathematics, the sciences, or a similarly unforgiving quantitative field. We care how you think, not your titles.
  • Strong mathematical foundations: linear algebra, probability, optimization, information theory
  • Can read a paper, decide what matters, and implement it
  • Have written real code for real computational systems; fluency with PyTorch and the modern ML stack, or the track record that says you will have it in weeks
  • Drawn to interpretability, model internals, and making systems provably reliable rather than usually fine
  • Self-directed, and able to make real progress without constant scaffolding

Our Stack

  • Languages: Python, Rust, and Node/TypeScript, with React on the frontend

  • Data: PostgreSQL, vector, and graph databases

  • Infra: Docker, Kubernetes, Terraform, across several cloud providers and customer VPCs

  • ML: Self-hosted models on multiple GPU providers and frontier APIs

Logistics

  • Full-time, in person in San Francisco
  • 10 to 12 weeks between May/June and August/September 2027
  • We sponsor US visas

What We Offer

World-Class Backing: You will join a venture-backed company with institutional investors including Google's Gradient Ventures, General Catalyst, and Y Combinator.

Real Impact: You will work directly on the core systems that determine how models perform in the wild. Your work ships into real, high-stakes environments where governance, auditability, and performance are non-negotiable.

Autonomy & Trust: We operate with a high degree of trust. You are expected to form strong technical opinions and execute on them.

To Apply

Apply through Work at a Startup with the single most interesting thing you have built or proven, such as a paper, a repo, or a calculation, and a sentence or two on why it mattered. We read every submission, this artifact is the signal.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
411,234 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$27k – $70k per year (Estimated) • Remote/Hybrid • 11+ years exp • Noida
JavaScript
Node JS
Python
SQL
TypeScript
Java
Java
Hibernate
Spring MVC
Databases
PostgreSQL
Frontend
JQuery
React.js
DevOps
Amazon CloudWatch
Amazon S3
AWS
AWS Lambda
CI/CD
Docker
IAM
Jenkins
Kubernetes
Platform Engineering
Rest API
Apply
DevOps Engineer II 1 day ago
In office • 5+ years exp • Bachelor's Degree
Go
JavaScript
Python
Ruby
DevOps
AWS
CI/CD
Configuration Management
Docker
Kubernetes
Platform Engineering
Apply
In office • Internship • Master's Degree • Paris
AI/ML
ChatGPT
Claude
Apply
In office • Internship • Master's Degree
AI/ML
ChatGPT
Claude
Copilot
Apply
$111k – $178k per year (Estimated) • Remote/Hybrid • Master's Degree • London
AI/ML
ChatGPT
Claude
Apply
$175k – $250k per year • Equity 0.5–1% • In office • Full-Time • 3+ years exp • San Francisco
AI/ML
LLM
Apply
$175k – $250k per year • Equity 0.5–1% • In office • Full-Time • 3+ years exp • San Francisco
AI/ML
Claude
Gemini
LLM
Apply
$96k – $144k per year • In office • Internship • Bachelor's Degree • San Francisco
JavaScript
Python
AI/ML
Claude
Gemini
LLM
DevOps
AWS
GCP
Git
Apply
$125k – $294k per year (Estimated) • Equity 0.5–1% • In office • Full-Time • 1+ year exp • San Francisco
AI/ML
LLM
PyTorch
Transformers
Interpretability
Apply
$175k – $250k per year • Equity 0.5–1% • In office • Full-Time • 3+ years exp • San Francisco
AI/ML
LLM
DevOps
Terraform
Apply
$85k – $105k per year • Equity 0–0.1% • In office • Full-Time • San Francisco
Management
Slack
Marketing
HubSpot
LinkedIn
Apply
$260k – $310k per year • Equity 0.1–0.4% • In office • Full-Time • 3+ years exp • San Francisco
Management
Slack
Marketing
HubSpot
Apply
In office • Internship • San Francisco
AI/ML
LLM
Management
Slack
Apply
$65k – $100k per year • Remote • Contractor • San Francisco
AI/ML
Claude
Apply
$71k – $95k per year • In office • 2+ years exp • Bachelor's Degree • San Francisco
Apply
See all jobs
This is one of many
411,234 more open roles from verified company boards, updated every day.