677,508open jobs
39,258companies
99,621added this week
Browse all
Salary
$104k – $258k per year (Estimated)
Location
Remote/Hybrid (Ramat Gan, Israel)
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
Grove Ventures is an Israeli early-stage venture capital fund partnering with seed and early-stage founders. It invests in deep tech and frontier technologies such as semiconductors, AI, quantum computing and digital health, often backing technical founding teams. The firm was founded in 2016 by Dov Moran, inventor of the USB flash drive, and is headquartered in Tel Aviv.

Description

We are seeking a Senior AI Researcher to lead post-training evaluation, red-teaming, and reinforcement learning (RL) gym audits on open-weight models. The ideal candidate will establish rigorous benchmarking methodologies, evaluate large language models (LLMs) against complex threats like Indirect Prompt Injections (IPI), and construct post-training evaluation pipelines that accurately measure realistic frontier-level security capabilities.

Key Responsibilities

  • RL Post-Training & Benchmarking: Execute post-training runs (e.g. GRPO) using mainstream open-weight generalist models against security-focused RL environments, targeting threat vectors like Indirect Prompt Injection (IPI). Reward Diagnostics & Trace Analysis - Analyze live loss curves and rollout traces to identify reward hacking, lazy policy convergence, and flawed or over/under-specified verifiers.
  • Task & Environment Auditing: Review tasks and multi-turn environments (including tool use, web navigation, and computer use) for realism, threat model accuracy, data distribution, and dataset balance.
  • Performance Reporting (Gym Cards): Generate comprehensive evaluation cards detailing hill-climbing performance uplift across checkpoints, failure modes, tokens/turns per rollout, and task-level success rates.
  • Integration & Orchestration: Integrate dockerized environments (e.g., Harbor format) into internal training frameworks, optimizing reset/statefulness semantics, concurrency, and throughput ceilings.

Requirements

Required Qualifications

  • Technical Background: M.S. or Ph.D. in Data Science, Machine Learning, Computer Science, or equivalent practical experience in deep learning.
  • RL & Post-Training Expertise: Strong hands-on experience training large-scale models using RL algorithms (e.g. GRPO, PPO) on open-weight architectures.
  • AI Security Expertise: Solid understanding of LLM vulnerabilities, red-teaming methodologies, and defensive alignment against IPI attacks.
  • Infrastructure Skills: Proficiency in PyTorch, Docker containerization, and distributed training architectures.
  • Diagnostic Skills: Ability to analyze agent rollout traces, craft deterministic rubrics/verifiers, and debug complex reward shaping flaws.

Preferred Qualifications

  • Prior experience working with standard RL gym formats, such as Harbor.
  • Experience evaluating complex agentic workflows in tool-use or web-browser environments.
  • Familiarity with evaluating open-weight models similar to Llama or Mistral against adversarial workloads.

About Alice

Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interact- whether with each other or with machines.

In a world where AI has fundamentally changed the nature of risk, Alice provides end-to-end coverage across the entire AI lifecycle. We support frontier model labs, enterprises, and UGC platforms with a comprehensive suite of solutions: from model hardening evaluations and pre-deployment red-teaming to runtime guardrails and ongoing drift detection.

Alice is widely considered a global leader in online safety and AI security. We have some of the most forward-thinking and passionate minds in the world working to safeguard over 3 billion users across the largest AI and tech platforms.

If you're creative and driven to secure the future of AI, we want to hear from you!

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
677,508 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Ramat Gan
$152k – $242k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • Santa Clara • Austin • Redmond
C++
C++
TensorFlow C++
PyTorch C++
LLVM
AI/ML
CUDA Toolkit
JAX
OpenCL
TensorFlow
PyTorch
CUDA
Triton
OpenAI
MLIR
Apache TVM
XLA
Apply
$152k per year • In office • 5+ years exp • Bachelor's Degree
Python
SQL
Python
pySpark
Databases
Apache Kafka
AI/ML
Hadoop
Spark
DevOps
CI/CD
Docker
Analytics
ETL/ELT
Apply
$153k – $180k per year • In office • 5+ years exp
Python
Databases
Databricks
AI/ML
LangChain
Spark
MLFlow
SHAP
Fine-tuning
Scikit-learn
Arize Phoenix
LIME
TensorFlow
PyTorch
OpenAI
Hugging Face
Amazon SageMaker
DevOps
GCP
Azure
AWS
Vector
Apply
Business Manager 4 hours ago
$140k per year • Remote/Hybrid • Full-Time • Bachelor's Degree • Irvine
SQL
AI/ML
AI Agents
LLM
Management
Confluence
Jira
Marketing
Mixpanel
Apply
$113k – $228k per year (Estimated) • Remote • Full-Time • 3+ years exp • Master's Degree • New York
AI/ML
vLLM
Function Calling
AI Agents
DPO
SFT
GRPO
Post-training
LLM Evaluation
LLM Guardrails
Tool Use
Apply
$85k – $168k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • London
Cybersecurity
VirusTotal
Recorded Future
Apply
$113k – $228k per year (Estimated) • Remote • Full-Time • 3+ years exp • Master's Degree • New York
AI/ML
vLLM
Function Calling
AI Agents
DPO
SFT
GRPO
Post-training
LLM Evaluation
LLM Guardrails
Tool Use
Apply
$113k – $228k per year (Estimated) • Remote • Full-Time • 3+ years exp • Master's Degree • San Francisco
AI/ML
vLLM
Function Calling
AI Agents
DPO
SFT
GRPO
Post-training
LLM Evaluation
LLM Guardrails
Tool Use
Apply
$113k – $228k per year (Estimated) • Remote • Full-Time • 3+ years exp • Master's Degree
AI/ML
vLLM
Function Calling
AI Agents
DPO
SFT
GRPO
Post-training
LLM Evaluation
LLM Guardrails
Tool Use
Apply
$56k – $134k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Ramat Gan
Apply
$112k – $254k per year (Estimated) • In office • Full-Time • 10+ years exp • Ramat Gan • Istanbul • Budapest • Lisbon
Cybersecurity
MITRE ATT&CK
Apply
$88k – $222k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • 7+ years exp • Bachelor's Degree • Ramat Gan
Python
Go
JavaScript
Node JS
Databases
Redis
Apache Kafka
DevOps
Azure
AWS
Cybersecurity
Crowdstrike
Apply
Senior Data Analyst 3 days ago
$75k – $171k per year (Estimated) • Remote • 5+ years exp • Master's Degree • Ramat Gan
Python
Python
pySpark
Databases
Databricks
AI/ML
Copilot
Cursor
Spark
Claude Code
DevOps
Terraform
Git
Apply
$118k – $219k per year (Estimated) • In office • Full-Time • 4+ years exp • Ramat Gan
Management
Agile
Apply
In office • Ramat Gan
Apply
See all jobs
This is one of many
677,508 more open roles from verified company boards, updated every day.