430,068open jobs
14,667companies
59,611added this week
Browse all
Salary
$64k – $190k per year
Location
In office (Seattle)
Employment
Full-Time
Overview
Company
Impact
Profile match
Handshake is a professional career network for students and early-career professionals headquartered in San Francisco, California, and founded in 2014. The company provides a digital platform that connects university students with internships and entry-level jobs, including specialized roles in the artificial intelligence economy. It serves over 15 million students and alumni from more than 1,500 colleges and universities, partnering with over 900,000 employers across the United States, United Kingdom, and Europe.

AI Red Teamer (LLM Generalist)

Location: Seattle, WA (candidates must reside in the Seattle metro area or be willing to relocate prior to start)

Type: Contract, 40 hours per week

About the Role

As an AI Red Teamer, you will stress-test large language models by intentionally trying to break them. Rather than checking whether an answer is correct, you will design creative, adversarial prompts that expose vulnerabilities: unsafe content, bias, broken guardrails, hallucinations, prompt injection weaknesses, and unexpected behaviors. Your work directly supports AI safety and model robustness for leading research labs.

This is a generalist red teaming role. You will probe models across the full spectrum of risk categories, including content safety, CBRN (chemical, biological, radiological, nuclear), cybersecurity, persuasion and influence operations, child safety, self-harm, over-companionship, and regulatory compliance. Red teaming may span text, image, voice, and agentic model capabilities depending on project needs.

This role requires creativity, curiosity, and an ability to think like an adversary while operating with strong ethical judgment.

Day-to-Day Responsibilities

  • Craft creative prompts and multi-turn scenarios to stress-test AI guardrails across diverse risk categories

  • Discover ways around safety filters, restrictions, and defenses using jailbreak, evasion, and prompt injection techniques

  • Explore edge cases to provoke disallowed, harmful, or incorrect outputs

  • Evaluate and score model responses against structured harm taxonomies and severity rubrics

  • Document experiments clearly, including what you tried, why you tried it, and what it revealed

  • Review and refine adversarial prompts generated by other team members

  • Contribute to harm taxonomy development, calibration exercises, and inter-rater reliability work

  • Collaborate with engineers, data scientists, and researchers to share findings and strengthen defenses

  • Work with potentially disturbing content on a regular basis (see Content Warning below)

  • Stay current on jailbreaks, attack methods, and evolving model behaviors

Desired Capabilities

Core

  • Strong hands-on experience using multiple LLMs (ChatGPT, Claude, Gemini, open-source models, etc.)

  • Intuition for crafting adversarial prompts; familiarity with jailbreak or evasion techniques is a strong plus

  • Creative, adversarial problem-solving skills

  • Clear and thoughtful written communication

  • Strong ethical judgment and the ability to separate adversarial thinking from personal values

  • Self-directed, collaborative, and comfortable in feedback-heavy environments

  • Curiosity, persistence, and comfort with frequent failure in experimentation

Nice to Have

  • Familiarity with Python or other scripting languages

  • Experience working with LLM APIs or evaluation tooling

  • Comfort with structured data annotation and rubric-based scoring

  • Prior work in trust and safety, content moderation, QA, or security research

  • Subject matter expertise in any high-risk domain (cybersecurity, chemistry, biology, medicine, law, finance, etc.)

You Will Thrive Here If

  • You treat every model response as a hypothesis to challenge

  • You can switch between creative free-association and rigorous documentation in the same session

  • You go deep into unusual interests (fandoms, niche internet cultures, gaming exploits, Wikipedia rabbit holes, etc.)

  • You come from a creative background: writing, visual art, improv, puzzle design, or similar

  • You are energized by finding the thing nobody else thought to try

  • You are genuinely passionate about AI and follow the space closely

Content Warning

This role involves regular and deliberate exposure to harmful content. You will encounter and intentionally generate content involving violence, self-harm, hate speech, sexually explicit material, child safety scenarios, and other categories of harmful output as part of structured adversarial testing. Candidates must be able to engage with this material professionally and sustainably. Support resources are available.

About Handshake AI

Handshake AI partners with leading AI research labs to make models safer and more robust. Our red teaming operations help identify vulnerabilities before they reach users, contributing directly to the responsible development of frontier AI systems.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
430,068 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Seattle
$60k – $111k per year (Estimated) • Remote • Full-Time • 4+ years exp
AI/ML
Claude
ChatGPT
DevOps
Incident Management
SLI/SLO/SLA
Apply
In office • Bachelor's Degree • Bengaluru
SQL
AI/ML
AI Agents
DevOps
Rest API
Apply
$36k – $95k per year (Estimated) • Remote • Full-Time
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
Tailwind CSS
Next.js
React.js
Radix UI
shadcn/ui
Framer Motion
DevOps
GitHub
Design
Framer
Apply
Remote • Full-Time
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
Tailwind CSS
Next.js
React.js
Radix UI
shadcn/ui
Framer Motion
DevOps
GitHub
Design
Framer
Apply
$73k – $194k per year (Estimated) • Remote • Full-Time
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
Tailwind CSS
Next.js
React.js
Radix UI
shadcn/ui
Framer Motion
DevOps
GitHub
Design
Framer
Apply
$88k – $115k per year • Remote • Full-Time • 1+ year exp
AI/ML
Scale AI
Apply
$136k – $170k per year • In office • Full-Time • 2+ years exp • San Francisco
AI/ML
Model Context Protocol
AI Agents
Post-training
Scale AI
Recommender Systems
Apply
$225k – $250k per year • In office • Full-Time • 8+ years exp • San Francisco
Python
JavaScript
AI/ML
LLM
Post-training
Scale AI
Knowledge Graph
DevOps
GCP
Azure
AWS
Apply
$39k – $85k per year (Estimated) • Equity • In office • Full-Time • Bachelor's Degree
Python
SQL
AI/ML
Post-training
Scale AI
Apply
$28k – $65k per year (Estimated) • Equity • In office • Full-Time • 6+ years exp
Python
SQL
Databases
Snowflake
Databricks
Google BigQuery
BigQuery
AI/ML
Post-training
Scale AI
Analytics
Tableau
Power BI
Looker
Apply
$60k – $118k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Chicago • Minneapolis • Dallas • Boston • Philadelphia
Robotics
Digital Twin
Analytics
Tableau
Power BI
Apply
$180k – $270k per year • In office • Full-Time • 10+ years exp • Seattle
Chips/EDA
Altium Designer
Apply
$145k – $163k per year • Remote • 4+ years exp • Bachelor's Degree • Seattle
Python
Go
JavaScript
TypeScript
SQL
Databases
PostgreSQL
AI/ML
AI Agents
Frontend
npm
DevOps
Rest API
Terraform
GCP
Packer
Azure
Git
AWS
Docker
Kubernetes
Amazon EKS
Amazon EC2
Amazon S3
IAM
Cryptography
Vault
Apply
$130k – $160k per year • Equity • Remote/Hybrid • 5+ years exp • Master's Degree • Seattle
Python
SQL
Python
pySpark
AI/ML
Spark
Apply
$163k – $386k per year (Estimated) • Remote/Hybrid • 8+ years exp • Seattle
Java
Ruby
DevOps
AWS
Management
Stripe
Apply
See all jobs
This is one of many
430,068 more open roles from verified company boards, updated every day.