Salary
$180k – $240k per year
Location
Remote (United States, United Kingdom, Canada)
Employment
Part-Time
Overview
Company
Impact
Profile match
Weekday is an Indian recruitment company that sources software engineers through referrals from other engineers rather than through job advertisements or agency databases. Its model pays working engineers to vouch for former colleagues they rate, turning informal knowledge about who is genuinely good into a searchable candidate pool that companies can hire from. Based in Bengaluru and backed by Y Combinator, the platform has layered AI screening and outbound sourcing on top of that referral network, and sells to startups and technology companies hiring in the Indian market.
This role is for one of our clients
Compensation: $90-$120 per hour
Join a leading AI lab's cutting-edge GenAI team and help build foundational AI models from the ground up. We're seeking MLOps Engineers with hands-on experience in large language model infrastructure across any of four areas: GPU kernel programming, performance profiling and trace analysis, debugging accelerated and distributed workloads, and high-throughput inference serving. This role involves AI model training and evaluation work, including writing and assessing MLOps and ML systems tasks and solutions to generate high-quality training data for frontier AI systems.
Requirements
Key Responsibilities
- Design challenging, domain-relevant tasks across four areas, GPU kernels, performance profiling, debugging, and inference serving, and write accurate, well-structured solutions to them.
- Guide research and engineering teams to close knowledge gaps and improve AI model performance on ML systems, training infrastructure, and framework-level topics.
- Evaluate MLOps and ML systems tasks and solutions, and provide clear, written technical feedback that stands up to reviewer scrutiny.
- Develop guidelines and detailed rubrics or evaluation frameworks covering kernel-level optimization, profiler output interpretation, distributed systems reasoning, and serving throughput and latency trade-offs.
- Collaborate with other subject matter experts to keep training data consistent and accurate.
Core Qualifications
- 2+ years of hands-on professional experience in ML systems, ML infrastructure, model serving, or GPU and accelerator performance engineering. This is a hands-on systems role rather than an applied modelling or data science one.
- Practical experience in at least one of the following, with more than one a strong plus: writing or optimizing custom GPU kernels (CUDA, Triton, Pallas); performance profiling and trace analysis (Kineto, torch.profiler, Nsight, XLA or JAX profiler); debugging distributed or accelerator-bound workloads; serving large language models at scale (vLLM, SGLang, TensorRT-LLM, Ray Serve, KV cache, paged attention, continuous batching).
- Working production experience with JAX and/or PyTorch. Framework-level depth is a strong plus: custom operators, distributed training (FSDP, DDP, DeepSpeed, Megatron), or compiler and graph-level work.
- Familiarity with modern accelerators such as A100, H100, B200 or TPU, and the ability to reason about throughput, latency and memory trade-offs.
- Demonstrable career progression.
- Ability to engage reliably for at least 40 hours/week during weekdays.
- Strong written communication skills and the ability to explain complex technical decisions clearly.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
682,105 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Free forever. No card. Under a minute.
Your match
How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.
Recommended for you based on this role
Similar stack
Same company
In your city
Research, Tinker, RL Systems
1 day ago
$350k – $475k per year • Remote/Hybrid • Full-Time • PhD • San Francisco
Python
Clarity
AI/ML
vLLM
Fine-tuning
Quantization
JAX
SGLang
TensorFlow
PyTorch
LLM
Post-training
Apply
$152k – $242k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • Santa Clara • Austin • Redmond
C++
C++
TensorFlow C++
PyTorch C++
LLVM
AI/ML
CUDA Toolkit
JAX
OpenCL
TensorFlow
PyTorch
CUDA
Triton
OpenAI
MLIR
Apache TVM
XLA
Apply
Senior Java EE Developer - AWS
4 hours ago
≈ $21k – $53k per year (Estimated) • In office • Full-Time • 7+ years exp • Johannesburg
Java
SQL
Apex
Java
Spring Boot
Hibernate
Hazelcast
Apache Camel
Apex
MuleSoft
Databases
MySQL
PostgreSQL
Redis
Oracle
RabbitMQ
Apache Kafka
AI/ML
JAX
DevOps
Rest API
Splunk
GCP
OpenShift
Azure DevOps
GitHub Actions
Prometheus
GitLab CI
Azure
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
Grafana
AWS Lambda
Amazon EC2
API Gateway
Cybersecurity
Keycloak
Apply
SDK SW Engineer
7 hours ago
≈ $33k – $72k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Bucharest
Python
C++
Cython
Cython
PyBind11
AI/ML
CUDA Toolkit
CUDA
Physical AI
Robotics
ROS
Apply
Member of Technical Staff, Agentic AI Architect
7 hours ago
≈ $146k – $332k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Singapore
Rust
TypeScript
AI/ML
LangGraph
LangChain
Model Context Protocol
vLLM
Embeddings
Scikit-learn
Prompt Engineering
Function Calling
Computer Vision
Chain-of-Thought
AI Agents
LangSmith
TensorRT
TensorRT-LLM
TensorFlow
PyTorch
CrewAI
LLM
RAG
A2A
LLM Evaluation
Multi-Agent Systems
Tool Use
DevOps
OpenTelemetry
CI/CD
Kubernetes
Vector
Apply
Apply
Personalized Life Assistant Expert
4 hours ago
$100k – $400k per year • Remote • Contractor
AI/ML
Cursor
Windsurf
Claude
ChatGPT
AI Agents
Gemini
LLM
Perplexity
OpenAI Codex
Apply
Apply
Apply
$300k – $500k per year • Remote • Contractor
Apply
This is one of many
682,105 more open roles from verified company boards, updated every day.

