725,615open jobs
43,268companies
103,521added this week
Browse all
Salary
$169k – $293k per year (Estimated)
Location
Remote (United States)
Seniority
Principal · 6+ years exp
Employment
Full-Time

First seen by Alion on Sep 24, 2026.

Overview
Company
Impact
Profile match
Decode and design the interfaces of life

This a Full Remote job, the offer is available from: New York (USA)

Principal ML Performance Engineer (GPU Optimization)

About Proxima

Proxima is a frontier AI and data generation company discovering the next generation of proximity therapeutics by making protein interactions programmable. Our platform brings together foundation-model machine learning, a scalable data generation engine, and a partnership track record exceeding $5B in collaborations across the world’s leading biopharma and tech organizations. We’ve recently closed an oversubscribed seed round with an elite group of VCs including DCVC, NVIDIA’s NVentures, AIX, Yosemite among others.

Neo-1 is our all-atom foundation model that combines state-of-the-art structure prediction and molecular generation in a single system. Neo-1 enables rapid exploration of chemical and structural space for high value, previously intractable targets, and in particular unlocks small molecule proximity therapeutics like molecular glues with AI for the first time.

In parallel, we are developing an advanced structural interactomics platform built on proprietary XLMS technology and a lab equipped with next-generation mass spectrometry instrumentation. This platform produces proteome-scale maps of protein interactions and helps identify small molecules that modulate proximity. Together with Neo-1, it creates an integrated system capable of co-folding protein complexes while generating candidate small molecules to influence those interactions.

Proximity-based therapeutics represent one of the most promising frontiers in modern drug discovery with the potential to treat previously intractable diseases and target ‘undruggable’ proteins. We’re building the tech and the team to make that happen. Come join us!

What you'll do

  • Profile and optimize training and inference for structural and generative models, including transformers, diffusion, and geometric deep learning

  • Write and tune custom kernels (CUDA, Triton) and use compilers (torch.compile, TensorRT, XLA) when beneficial

  • Scale distributed training across 32-64 nodes, employing FSDP, DeepSpeed, tensor and pipeline parallelism, and mixed precision

  • Reduce inference cost by optimizing memory scaling for large complexes, improving diffusion sampling efficiency, batching ragged inputs, and maximizing throughput across up to 1000 GPUs

  • Manage GPU cluster efficiency on GCP, focusing on scheduling, utilization, spot strategy, and cost reporting

  • Develop benchmarks and profiling tools for the research team

What we need

  • Minimum of 6+ years experience in ML systems, HPC, or performance engineering, with a BS/MS/PhD in CS, EE, or related field

  • Demonstrated ability to set technical direction beyond coding: selecting infrastructure, influencing research teams, and mentoring engineers

  • Deep knowledge of PyTorch internals with hands-on experience profiling and fixing real bottlenecks

  • Experience with CUDA and Triton, skilled at reading Nsight output, and strong understanding of memory bandwidth and occupancy

  • Experience with distributed training at multi-node scale

  • Strong proficiency in Python and C++

  • Able to name a model they made materially faster and quantify the improvement

Nice to haves

  • Experience in geometric deep learning, equivariant networks, or protein structure models such as AlphaFold, ESM, or RFdiffusion

  • Experience writing kernels for structure-model primitives, including triangle attention, triangle multiplicative updates, cuEquivariance, or FlashAttention for pair bias

  • Experience orchestrating large batch inference and managing Kubernetes GPU scheduling

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
725,615 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
New York
$18k per year (net) • In office • Kazan
Python
C++
C++
CMake
PyTorch C++
AI/ML
Transformers
PyTorch
LLM
Hugging Face
Embodied AI
DevOps
Rest API
WebSockets
Git
Docker
Linux
Robotics
ROS
Apply
$18k per year (net) • In office • Saint Petersburg
Python
C++
C++
CMake
PyTorch C++
AI/ML
Transformers
PyTorch
LLM
Hugging Face
Embodied AI
DevOps
Rest API
WebSockets
Git
Docker
Linux
Robotics
ROS
Apply
$18k per year (net) • In office • Krasnodar
Python
C++
C++
CMake
PyTorch C++
AI/ML
Transformers
PyTorch
LLM
Hugging Face
Embodied AI
DevOps
Rest API
WebSockets
Git
Docker
Linux
Robotics
ROS
Apply
Senior Data Engineer 2 hours ago
$27k – $57k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
AI/ML
Spark
MLFlow
DevOps
Terraform
GitHub Actions
CI/CD
AWS
Platform Engineering
FinOps
Incident Management
Amazon S3
IAM
Apply
$9k – $23k per year (Estimated) • In office • 1+ year exp • Bachelor's Degree • Tashkent
Python
SQL
Analytics
Power BI
Apply
$106k – $209k per year (Estimated) • In office • 4+ years exp • Master's Degree • New York
Apply
$80k – $120k per year • In office • Full-Time • 1+ year exp • New York
AI/ML
RAG
OpenAI
Management
n8n
Marketing
Reddit
Apply
$150k – $300k per year • Equity • Remote • Full-Time • 3+ years exp • New York
AI/ML
RAG
DevOps
Grafana
Management
Monday.com
n8n
Apply
GTM Associate (US) 1 day ago
$70k – $120k per year • In office • Full-Time • 1+ year exp • New York
AI/ML
RAG
OpenAI
DevOps
Grafana
Marketing
Reddit
Apply
$70k – $110k per year • In office • Full-Time • 1+ year exp • New York
AI/ML
RAG
OpenAI
DevOps
Docker
Marketing
Reddit
Apply
See all jobs
This is one of many
725,615 more open roles from verified company boards, updated every day.