745,032open jobs
44,747companies
107,631added this week
Browse all
Salary
$200k per year
Location
Hybrid (New York, United States)
Seniority
Middle · 3+ years exp

Confirmed on the employer's own hiring board on Sep 24, 2026. First seen by Alion on Sep 24, 2026. Tower Research Capital scores B on the Alion truth index.

Overview
Company
Impact
Profile match
Tower Research Capital is a quantitative trading firm operating automated strategies globally. It builds its own trading infrastructure, exchange connectivity and research systems. The firm supports multiple independent trading teams under one platform.

Tower Research Capital is a leading quantitative trading firm founded in 1998. Tower has built its business on a high-performance platform and independent trading teams. We have a 25+ year track record of innovation and a reputation for discovering unique market opportunities.

Tower is home to some of the world’s best systematic trading and engineering talent. We empower portfolio managers to build their teams and strategies independently while providing the economies of scale that come from a large, global organization. 

Engineers thrive at Tower while developing electronic trading infrastructure at a world class level. Our engineers solve challenging problems in the realms of low-latency programming, FPGA technology, hardware acceleration and machine learning. Our ongoing investment in top engineering talent and technology ensures our platform remains unmatched in terms of functionality, scalability and performance.

At Tower, every employee plays a role in our success. Our Business Support teams are essential to building and maintaining the platform that powers everything we do - combining market access, data, compute, and research infrastructure with risk management, compliance, and a full suite of business services. Our Business Support teams enable our trading and engineering teams to perform at their best.

At Tower, employees will find a stimulating, results-oriented environment where highly intelligent and motivated colleagues inspire each other to reach their greatest potential.

Summary

You will bridge the gap between quantitative research and high-performance computing, building and optimizing the systems used to train machine learning models at scale. You will focus on accelerating the end-to-end training lifecycle-from data ingestion and distributed execution to kernel performance and hardware utilization-enabling researchers to iterate more quickly across increasingly complex models and datasets.

Responsibilities

  • Training Performance and Benchmarking
    • Benchmark model-training workloads across CPUs, GPUs, and other accelerator platforms to identify bottlenecks and guide Tower’s compute infrastructure decisions.
    • Develop performance models and standardized benchmarks for measuring throughput, utilization, scalability, and time to convergence.
  • Distributed Training Optimization
    • Design and optimize distributed training strategies, including data, tensor, pipeline, and model parallelism.
    • Improve communication efficiency across multi-GPU and multi-node environments by optimizing collective operations, topology awareness, and computation-communication overlap.
  • End-to-End Training Efficiency
    • Analyze and improve the full training pipeline, including data loading, preprocessing, memory management, forward and backward passes, optimizer execution, checkpointing, and experiment recovery.
    • Identify bottlenecks across compute, memory, storage, networking, and interconnects to increase accelerator utilization and researcher productivity.
  • GPU Kernel and Framework Development
    • Develop and optimize GPU kernels and performance-critical framework components for quantitative machine learning workloads.
    • Integrate specialized libraries, compilers, and execution techniques to improve throughput, memory efficiency, and numerical performance.
  • Model and Numerical Optimization
    • Apply techniques such as mixed-precision training, gradient accumulation, activation checkpointing, operator fusion, and memory-efficient optimizers.
    • Evaluate tradeoffs among training speed, numerical stability, reproducibility, model quality, and infrastructure cost.
  • Training Infrastructure
    • Partner with HPC and infrastructure teams to optimize workload scheduling, resource allocation, observability, fault tolerance, and reproducibility across shared compute environments.
    • Help define the architecture and tooling required to support large-scale experimentation across on-premises and cloud-based infrastructure.
  • Cross-Functional Collaboration
    • Work closely with ML Researchers, Quantitative Researchers, HPC Engineers, Systems Engineers, and hardware specialists to translate research requirements into highly efficient training systems.

Qualifications

  • 3+ years of experience optimizing machine learning training workloads in high-performance, distributed, or large-scale computing environments.
  • Deep knowledge of machine learning frameworks such as PyTorch or JAX, including their execution models, compilation paths, autograd systems, and distributed-training capabilities.
  • Strong programming skills in Python and C++, with experience developing or optimizing performance-critical systems.
  • Proven experience with GPU kernel development and optimization using technologies such as CUDA, Triton, CUTLASS, cuBLAS, cuDNN, or related libraries.
  • Strong understanding of GPU architecture, including streaming multiprocessor execution, warp scheduling, tensor cores, and the memory hierarchy from registers through HBM.
  • Experience with distributed-training technologies and communication libraries such as NCCL, FSDP, DeepSpeed, Megatron-LM, XLA, or equivalent systems.
  • Proficiency with performance-analysis tools such as Nsight Systems, Nsight Compute, PyTorch Profiler, or comparable tracing and profiling platforms.
  • Understanding of high-performance networking, storage, and accelerator interconnects, including technologies such as InfiniBand, RDMA, NVLink, or NVSwitch.
  • Demonstrated ability to benchmark heterogeneous compute platforms and make rigorous, data-driven recommendations about performance, scalability, and cost.

Preferred Qualifications

  • Experience optimizing training workloads for transformer-based, time-series, reinforcement-learning, or other computationally intensive models.
  • Experience with cluster orchestration and scheduling technologies such as Kubernetes, Slurm, Ray, or similar platforms.
  • Familiarity with fault-tolerant distributed training, large-scale checkpointing, experiment reproducibility, and GPU-cluster observability.
  • Practical experience with specialized accelerators, custom hardware, or compiler technologies for machine learning.
  • Prior experience in financial trading is not required.

Anticipated New York annual base salary of $200,000, plus eligible for discretionary bonus.

Benefits

Tower’s headquarters are in the historic Equitable Building, right in the heart of NYC’s Financial District and our impact is global, with over a dozen offices around the world. 

At Tower, we believe work should be both challenging and enjoyable. That is why we foster a culture where smart, driven people thrive - without the egos. Our open concept workplace, casual dress code, and well-stocked kitchens reflect the value we place on a friendly, collaborative environment where everyone is respected, and great ideas win.

Our benefits include:

  • Generous paid time off policies
  • Savings plans and other financial wellness tools available in each region
  • Hybrid working opportunities
  • Free breakfast, lunch and snacks daily 
  • In-office wellness experiences and reimbursement for select wellness expenses (e.g., gym, personal training and more) 
  • Volunteer opportunities and charitable giving 
  • Social events, happy hours, treats and celebrations throughout the year
  • Workshops and continuous learning opportunities

At Tower, you’ll find a collaborative and welcoming culture, a diverse team and a workplace that values both performance and enjoyment. No unnecessary hierarchy. No ego. Just great people doing great work - together.

Tower Research Capital is an equal opportunity employer.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
745,032 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
New York
≈ $145k – $253k per year (Estimated) • Equity • Remote (United States) • Full-Time • 5+ years exp • United States
Python
Go
AI/ML
AI Agents
LLM
Knowledge Graph
Agentic Workflows
DevOps
Kubernetes
Platform Engineering
Cybersecurity
Crowdstrike
Analytics
A/B Testing
Apply
AI Software Engineer 11 days ago
$110k – $150k per year • Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • United States
Python
JavaScript
Java
TypeScript
Node JS
Python
FastAPI
Java
Spring Boot
AI/ML
LangChain
Fine-tuning
Prompt Engineering
Multimodal AI
AI Agents
LLM
RAG
OpenAI
Hugging Face
LLMOps
Edge AI
Machine Learning
Frontend
Angular
React.js
DevOps
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Apply
≈ $137k – $250k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • PhD • United States
Python
Java
Java
Spring Boot
AI/ML
Machine Learning
DevOps
GCP
GitHub Actions
Azure
CI/CD
Jenkins
AWS
Twelve-Factor App
Apply
CXO AI Engineer 3 hours ago
≈ $152k – $265k per year (Estimated) • Equity • Remote (United States) • 5+ years exp
Python
SQL
Databases
Amazon Redshift
Trino
AI/ML
Model Context Protocol
dbt
Function Calling
AI Agents
LLM
RAG
Context Engineering
Tool Use
DevOps
SLI/SLO/SLA
Management
Slack
Apply
$74k – $220k per year • Remote (United States) • Full-Time • 7+ years exp • High School Diploma • Philadelphia
Python
Java
AI/ML
Copilot
Cursor
LangGraph
AutoGen
LangChain
Claude Code
Prompt Engineering
AI Agents
Semantic Kernel
CrewAI
RAG
Multi-Agent Systems
DevOps
Rest API
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Apply
≈ $33k – $89k per year (Estimated) • In office • Internship
Python
JavaScript
TypeScript
Python
Flask
Django
Databases
MySQL
PostgreSQL
Frontend
Next.js
React.js
DevOps
Terraform
Azure DevOps
Azure
AWS
Docker
GitHub
Management
Slack
Apply
≈ $47k – $117k per year (Estimated) • In office • Internship
Python
AI/ML
PyTorch
DevOps
Linux
Apply
≈ $50k – $125k per year (Estimated) • In office • Tokyo
Python
AI/ML
PyTorch
DevOps
Linux
Apply
≈ $46k – $116k per year (Estimated) • In office • Freelance
Python
Apply
AIエンジニア 1 hour ago
≈ $47k – $118k per year (Estimated) • In office
Python
AI/ML
PyTorch
DevOps
Linux
Management
Slack
Apply
$120k – $200k per year • Hybrid • 2+ years exp • Bachelor's Degree • New York
AI/ML
LangChain
Model Context Protocol
AI Agents
Multi-Agent Systems
Machine Learning
Apply
GPU Systems Engineer 1 month ago
$200k – $300k per year • Hybrid • 5+ years exp • New York
Python
C++
AI/ML
CUDA Toolkit
CUDA
NCCL
NVLink
Machine Learning
DevOps
Puppet
Ansible
Chef
Configuration Management
Self-Healing
HPC
Linux
Apply
$200k – $300k per year • Hybrid • 2+ years exp • Bachelor's Degree • New York
Python
AI/ML
Machine Learning
DevOps
CI/CD
Linux
Unix
Apply
$200k – $300k per year • Hybrid • 2+ years exp • New York
Python
C++
C++
PyTorch C++
AI/ML
Quantization
JAX
Knowledge Distillation
ONNX
TensorRT
PyTorch
Triton
IREE
CUTLASS
Model Distillation
Machine Learning
DevOps
HPC
Apply
$200k – $300k per year • Hybrid • New York
Python
Python
Dask
AI/ML
AI Agents
TensorFlow
PyTorch
LLM
Ray
Time Series Forecasting
Agentic Workflows
Machine Learning
Apply
$81k – $92k per year • In office • Full-Time • 5+ years exp • High School Diploma • McLean • New York
Management
Outlook
Microsoft Office
Apply
$215k – $246k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • New York
Python
Java
SQL
Scala
Databases
Snowflake
Databricks
Cassandra
DynamoDB
Amazon Redshift
AI/ML
Spark
Dagster
Machine Learning
DevOps
Splunk
GCP
Azure
AWS
Management
Agile
Apply
$245k – $279k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • San Francisco • New York • San Jose
Python
Go
JavaScript
Rust
TypeScript
C#
Scala
AI/ML
AI Agents
Machine Learning
DevOps
GCP
Azure
AWS
HPC
Apply
$201k – $229k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • New York • McLean
Apply
$220k – $325k per year • Hybrid • Full-Time • 10+ years exp • New York
AI/ML
AI Agents
LLM
DevOps
CI/CD
Apply
See all jobs
This is one of many
745,032 more open roles from verified company boards, updated every day.