1,368,308open jobs
79,536companies
209,295added this week
Browse all
Salary
$120k – $180k per year
Location
Remote (Singapore)
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 8, 2026. First seen by Alion on Sep 1, 2026.

Overview
Company
Impact
Profile match
SASH builds AI safety research capacity in Asia and policy-relevant governance work from Singapore.

About the team

Neo Research (新衡) is an independent AI safety research organization based in Singapore. We study frontier risks in increasingly capable AI systems, with a particular focus on the open-weight model ecosystem and the rapidly growing frontier-model ecosystem in Asia.

Some of the world’s most capable open-weight models are now being developed in Asia and deployed globally. Yet, they remain poorly understood from a frontier-safety perspective. We want to understand how to ensure their safety, and what new risks become important as the frontier changes.

Our current work focuses on misalignment, loss of control, and harmful manipulation. We study questions such as whether models pursue unintended objectives, conceal problematic behaviour, recognize and adapt to evaluations, evade oversight, or become less safe as they are given greater autonomy. Our goal is to produce rigorous empirical evidence about risks that are important but difficult to measure.

We are looking for Research Engineers to build the experimental systems needed to evaluate increasingly capable models: realistic agent environments, model and tool integrations, long-horizon evaluation infrastructure, and the systems needed to make complex experiments reliable and reproducible.

Why join Neo Research

Evaluating increasingly capable models requires more than running benchmarks. It means building realistic environments, giving models tools and autonomy, running experiments that may unfold over hundreds of interactions, and capturing enough information to distinguish genuine model behaviour from quirks or failures of the evaluation itself.

You will work closely with researchers while owning substantial parts of the experimental system,from agent scaffolds and tool integrations to model sampling, observability, trajectory analysis, and reproducibility. In this role, you will go beyond implementing evaluations and actively help shape evaluation methodology.

Your work will contribute directly to published research shared with AI Safety Institutes, frontier labs, policymakers, and model developers. We have presented our work to most major Chinese model developers and run a joint evaluations project with an AI Safety Institute. Ourmission statement describes our current research directions.

In this role, you would:

  • Design, implement, and run evaluations of misalignment and loss-of-control risks in frontier models.

  • Build realistic agent environments, scaffolds, and tool integrations for studying behaviour under increasing autonomy.

  • Develop infrastructure for long-horizon experiments involving many model calls, actions, tools, and environment states.

  • Own experimental reliability and reproducibility across model access, sampling, configuration, environment management, logging, and analysis.

  • Work with research scientists to turn open-ended questions into tractable experiments, and identify cases where an apparent model behaviour is actually an artefact of the evaluation.

  • Build tools for inspecting and analysing large collections of model trajectories and comparing behaviour across models and experimental conditions.

  • Contribute experimental methodology, infrastructure, and technical analysis directly to published research.

About you

Essentials

  • Strong Python and general software-engineering skills.

  • Experience working with language models, agentic systems, or model evaluations.

  • The ability to turn an underspecified research question into a reliable experimental setup.

  • Strong engineering practices, particularly around reproducibility, testing, observability, and data integrity.

  • The ability to investigate unexpected results across both the model and the surrounding infrastructure.

  • Clear technical writing and communication skills.

  • Comfort working in an early-stage environment where requirements and research directions evolve.

Helpful, but not required

  • Experience with AI safety or dangerous-capability evaluations.

  • Experience with evaluation frameworks such as Inspect.

  • Experience building agent scaffolds or tool-use environments.

  • Experience with distributed inference, large-scale API-based experimentation, Docker, Kubernetes, or related infrastructure.

  • Experience analysing large collections of model trajectories or transcripts.

  • Familiarity with frontier-model safety reports.

  • Mandarin reading or writing ability.

You don't need to have worked in AI safety specifically. We're also interested in strong ML and research engineers whose technical expertise and engineering judgment could transfer strongly to this work.

Role logistics & benefits

Location:

This role can be based in Singapore or remotely.

Our team is globally distributed, so remote team members should be comfortable maintaining some working-hour overlap with colleagues across regions.

Compensation:

Our compensation takes location, experience, and level into account, with indicative salary ranges of $120,000-$180,000+. We may offer above this range for exceptional candidates.

Benefits:

Competitive benefits and leave policies.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,368,308 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
In your city
$100k – $150k per year • Remote (United States) • 6+ years exp • Bachelor's Degree
Python
AI/ML
Fine-tuning
LLM
LLM Guardrails
Machine Learning
Cybersecurity
Threat Modeling
Apply
$100k – $150k per year • Remote (United States) • 6+ years exp • Bachelor's Degree
Python
Go
C++
C++
PyTorch C++
AI/ML
DeepSpeed
JAX
PyTorch
Ray
Megatron-LM
FSDP
NCCL
InfiniBand
DevOps
SLURM
CI/CD
Kubernetes
FinOps
HPC
Linux
Apply
$100k – $150k per year • Remote (United States) • 6+ years exp • Bachelor's Degree
Python
C++
AI/ML
Quantization
Knowledge Distillation
Federated Learning
Edge AI
LiteRT
ONNX Runtime
Model Distillation
Machine Learning
Mobile
Core ML
Apply
$100k – $150k per year • Remote (United States) • 6+ years exp • Bachelor's Degree
Python
AI/ML
Spark
Multimodal AI
Ray
Human-in-the-Loop
DevOps
CI/CD
Apply
$166k – $228k per year • Remote (United States) • 8+ years exp • Master's Degree
Python
ABAP
ABAP
SAP BTP
Databases
PostgreSQL
Snowflake
Apache Kafka
Google BigQuery
Amazon Redshift
BigQuery
AI/ML
Cursor
Claude
Spark
MLFlow
Vertex AI
dbt
Prefect
TensorFlow
PyTorch
LLM
Hugging Face
Amazon SageMaker
Feature Store
DevOps
Terraform
GCP
GitHub Actions
CloudFormation
Pulumi
Azure
CI/CD
AWS
Docker
Kubernetes
AWS Lambda
Amazon EventBridge
Cybersecurity
SOC 2
GDPR
HIPAA
Analytics
ETL/ELT
Apply
≈ $51k – $115k per year (Estimated) • Remote (Argentina) • Full-Time • Argentina
Java
Kotlin
SQL
Java
Spring Boot
Gradle
Kotlin
Mockito
Databases
RabbitMQ
Apache Kafka
Mobile
JUnit
DevOps
Rest API
GitHub Actions
Kibana
Dynatrace
GitLab CI
CI/CD
Jenkins
AWS
Docker
Kubernetes
Robotics
Path Planning
Management
Agile
Apply
≈ $33k – $87k per year (Estimated) • Remote (Argentina) • Full-Time • 7+ years exp • Argentina
Python
SQL
Python
Pydantic
HTTPX
AI/ML
Copilot
Cursor
Claude Code
DevOps
Rest API
GitHub Actions
CircleCI
GitLab CI
CI/CD
Jenkins
AWS
Docker
Robotics
Path Planning
Analytics
ETL/ELT
Management
Agile
Scrum
QA
JMeter
Cypress
Pytest
k6
Locust
Apply
≈ $25k – $66k per year (Estimated) • Hybrid • Full-Time • Bachelor's Degree • Guadalajara
Databases
Oracle
AI/ML
AI Agents
Edge AI
DevOps
VMWare
Management
ServiceNow
Agile
ITIL
Apply
$200k – $275k per year • Remote (United States) • Full-Time • 12+ years exp • United States
SQL
AI/ML
AutoGen
LangChain
Fine-tuning
AI Agents
CrewAI
LLM
RAG
Agentic Workflows
Analytics
Alation
Apply
$103k – $160k per year • Remote (Canada) • Full-Time • 7+ years exp
AI/ML
AI Agents
LLM Evaluation
Agentic Workflows
Apply
$200k – $250k per year • Remote (Singapore) • Full-Time • PhD
AI/ML
AI Agents
EU AI Act
NIST AI RMF
Machine Learning
Apply
$150k – $200k per year • Remote (Singapore) • Full-Time • PhD
AI/ML
AI Agents
EU AI Act
NIST AI RMF
Apply
$90k – $150k per year • Remote (United Kingdom, Singapore) • Full-Time • London • San Francisco
Apply
$90k – $150k per year • In office • Full-Time • London
AI/ML
CUDA Toolkit
LLM
CUDA
Post-training
KV Cache
Apply
$90k – $150k per year • In office • Full-Time • London
Cybersecurity
PKI
Apply
See all jobs
This is one of many
1,368,308 more open roles from verified company boards, updated every day.