Megatron-LM Jobs - Remote & On-site

Megatron-LM jobs hiring now at vetted startups and product companies on Alion, updated daily. Average salary around $329,084/year. Compare remote, hybrid and on-site positions with relocation and visa sponsorship. Browse and apply today.

AI/ML
Data Science
Backend
Frontend
Mobile
DevOps
Web3
Games
Hardware
Robotics
Security
QA
Executive
Networking
Product
Design
Analytics
Support
Enterprise Apps
Quantum

Liquid AI

liquid.ai
Liquid AI is an artificial intelligence company headquartered in Boston, Massachusetts, and founded in 2023 as a spin-off from the MIT Computer Science and Artificial Intelligence Laboratory. The company builds Liquid Foundation Models, an architecture derived from liquid neural networks that aims to match transformer quality at a fraction of the memory and compute. It targets on-device and edge deployment where models must run on phones, vehicles, and embedded hardware rather than in a data center.
liquid.ai • HQ: Boston, United States • AI Infrastructure • Hardware • Machine Learning • Edge AI • LLM & Generative AI • Artificial Intelligence • Est. 2023
HQ: Boston, United States • AI Infrastructure • Hardware • Machine Learning • Edge AI • LLM & Generative AI • Artificial Intelligence • Est. 2023
Open 311 daysVerified live · 1 day ago 10 months ago

Member of Technical Staff - Multi-Modal, Vision

$194k – $394k per year (Estimated) • Staff+ • Remote/Hybrid (San Francisco, United States) • Master's Degree • Full-Time • Staff
Python
AI/ML
Computer Vision
DeepSpeed
Multimodal AI
Reinforcement Learning
VLM
FSDP
Hugging Face
Megatron-LM
Post-training
SFT
DevOps
GitHub
Apply
Open 397 daysVerified live · 1 day ago 1 year ago

Member of Technical Staff - Distributed Training Engineer

$199k – $404k per year (Estimated) • Staff+ • Remote/Hybrid (San Francisco, United States) • Full-Time • Staff
DeepSpeed
Multimodal AI
PyTorch
FSDP
Megatron-LM
NCCL
Mixture of Experts
Apply
Report

Hark

hark.com
Hark was an online digital entertainment platform best known for its extensive library of short audio soundbites, video clips, and pop culture quotes. Launched in 2007, the website allowed users to browse, create, and share playable soundboards featuring memorable lines from movies, television shows, and political figures. While it grew into a popular destination for viral sound clips during the late 2000s and early 2010s, the platform has since ceased its original operations.
hark.com • HQ: San Jose, United States • Machine Learning • Artificial Intelligence
HQ: San Jose, United States • Machine Learning • Artificial Intelligence
Open 160 daysTop 25% pay 5 months ago

Member of Technical Staff, Pretraining

$180k – $450k per year • Staff+ • In office (San Jose, United States) • Full-Time • Staff
DeepSpeed
LLM
Multimodal AI
Synthetic Data
Megatron-LM
AI Agents
Apply
Open 160 daysTop 25% pay 5 months ago

Member of Technical Staff, Multimodal Vision

$180k – $450k per year • Staff+ • In office (San Jose, United States) • Full-Time • Staff
DeepSpeed
Fine-tuning
Multimodal AI
Reinforcement Learning
RLHF
DPO
FSDP
GRPO
Megatron-LM
Post-training
PPO
SFT
AI Agents
Apply
Report

Perplexity AI

perplexity.ai
Perplexity AI is an American company that builds an answer engine combining live web search with large language models to return sourced, conversational responses instead of a list of links. Its products span a consumer assistant on web and mobile, the Comet browser, enterprise search over internal documents and the Sonar developer API that exposes the same grounded retrieval stack. Founded in 2022 in San Francisco by former researchers and engineers from OpenAI, Meta and Databricks, the company is backed by NVIDIA, IVP, New Enterprise Associates and SoftBank.
perplexity.ai • HQ: San Francisco, United States • AI Agents • Internet Services • Artificial Intelligence • Search Engines • LLM & Generative AI • 1001-5000 employees • Est. 2022
HQ: San Francisco, United States • AI Agents • Internet Services • Artificial Intelligence • Search Engines • LLM & Generative AI • 1001-5000 employees • Est. 2022
Open 139 daysTop 25% pay 4 months ago

Member of Technical Staff (AI Researcher)

$220k – $485k per year • Staff+ • 2+ years expIn office (San Francisco, Palo Alto, United States) • PhD • Full-Time • Staff
Python
C++
C++
PyTorch C++
AI/ML
Fine-tuning
LLM
Perplexity
PyTorch
Reinforcement Learning
DPO
GRPO
Megatron-LM
Post-training
SFT
CUDA Toolkit
CUDA
Apply
Report

SambaNova Systems

sambanova.ai
SambaNova Systems is an American artificial intelligence semiconductor and cloud platform company headquartered in Palo Alto, California. Founded in 2017 by Stanford professors Kunle Olukotun and Christopher Ré alongside former Oracle executive Rodrigo Liang, the company builds full-stack hardware and software infrastructure purpose-built for enterprise AI, large-scale model training, and high-speed agentic AI inference.
sambanova.ai • HQ: Palo Alto, United States • Artificial Intelligence • Semiconductors • AI Infrastructure
HQ: Palo Alto, United States • Artificial Intelligence • Semiconductors • AI Infrastructure
Open 282 daysVerified live · 1 hour ago 9 months ago

Senior AI Performance Engineer

$151k – $293k per year (Estimated) • Senior • 3+ years expIn office (San Jose, United States) • Bachelor's Degree • Full-Time • Senior
C++
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
DeepSeek
JAX
Llama
PyTorch
Quantization
Qwen
TensorFlow
CUDA
CUDA Toolkit
DeepSpeed
LLM
Multimodal AI
OpenCL
TensorRT
Triton
vLLM
cuDNN
Megatron-LM
Apply
Report

The True Engineer

thetrueengineer.com
The True Engineer is an engineering leadership and software development publication platform hosted on Substack. Authored by Adlet Balzhanov, a senior software engineering lead with background in Big Tech and high-growth enterprise scale-ups, the publication delivers monthly long-form analytical essays, career frameworks, and strategic insights for software developers, engineering managers, and technology executives. Operating under a creator-driven digital subscription model, the newsletter focuses on technical leadership tactics, staff-level engineering practices, organizational dynamics within large tech companies, and navigating the operational impacts of AI on software development.
thetrueengineer.com • HQ: Stockholm, Sweden • Artificial Intelligence
HQ: Stockholm, Sweden • Artificial Intelligence
Verified live · 2 hours ago 8 days ago

ML Systems Performance Engineer (MFU)

$22k – $54k per year (Estimated) • Equity • In office (Almaty, Kazakhstan) • Relocation • Full-Time
C
C
MPI
AI/ML
CUDA
CUDA Toolkit
DeepSpeed
PyTorch
Triton
FSDP
Megatron-LM
Multimodal AI
Reinforcement Learning
InfiniBand
NCCL
NVLink
Mixture of Experts
Apply
Report

NVIDIA

nvidia.com
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.
nvidia.com • HQ: Santa Clara, United States • AI Infrastructure • Semiconductors • AI Agents • Big Data • LLM & Generative AI • Machine Learning • Hardware • Data & Analytics • Information Technology • Artificial Intelligence • 5000+ employees • Est. 1993
HQ: Santa Clara, United States • AI Infrastructure • Semiconductors • AI Agents • Big Data • LLM & Generative AI • Machine Learning • Hardware • Data & Analytics • Information Technology • Artificial Intelligence • 5000+ employees • Est. 1993
$184k – $288k per year • Senior • 5+ years expIn office (Santa Clara, United States) • Master's Degree • Full-Time • Senior
C++
Python
C++
PyTorch C++
AI/ML
LLM
PyTorch
Ray
Reinforcement Learning
RLHF
AI Agents
DPO
FSDP
GRPO
Post-training
PPO
DeepSpeed
DeepSpeed-Chat
Multimodal AI
Quantization
SGLang
TensorRT
TensorRT-LLM
vLLM
VLM
InfiniBand
Megatron-LM
Mixture of Experts
NCCL
NVIDIA NeMo
NVLink
Pre-training
DevOps
Kubernetes
Apply
$152k – $242k per year • Senior • 4+ years expIn office (Redmond, Santa Clara, United States) • Master's Degree • Full-Time • Senior
C++
Python
C++
PyTorch C++
AI/ML
LLM
PyTorch
SGLang
Triton
vLLM
Hugging Face
Megatron-LM
Mixture of Experts
SFT
Apply
$44k – $105k per year (Estimated) • Staff+ • 7+ years expIn office (Bengaluru, Mumbai, India) • Master's Degree • Full-Time • Architect
Fine-tuning
LLM
PyTorch
RAG
AI Agents
Megatron-LM
DevOps
AWS
Azure
Docker
GCP
Kubernetes
Apply
Report

JetBrains

jetbrains.com
JetBrains is a global software vendor specializing in intelligent, productivity-enhancing developer tools and integrated development environments (IDEs). Founded in 2000 and headquartered in Prague, Czech Republic, the company is best known for creating IntelliJ IDEA, PyCharm, WebStorm, Rider, ReSharper, and the Kotlin programming language (which serves as Google's preferred language for Android development).
jetbrains.com • HQ: Prague, Czech Republic • Artificial Intelligence • Software • Code Intelligence • Productivity Software • 51-200 employees • Est. 2000
HQ: Prague, Czech Republic • Artificial Intelligence • Software • Code Intelligence • Productivity Software • 51-200 employees • Est. 2000
Verified live · 14 min ago 4 months ago

Senior Research Engineer (Agentic Behavior)

Senior • 3+ years expAmsterdam • Remote (Germany, Cyprus, Czech Republic, Netherlands, Poland) • Relocation • Senior
Java
Kotlin
Python
SQL
Java
Gradle
Kotlin
Ktor
AI/ML
AI Agents
Claude
Claude Code
Cursor
Langfuse
LLM
MLFlow
Promptfoo
PyTorch
RLHF
TRL
Transformers
Anthropic
Context Engineering
DPO
GRPO
Megatron-LM
OpenAI
Post-training
SFT
Model Context Protocol
Mobile
Kotlin Multiplatform
Apply
Verified live · 14 min ago 7 months ago

Research Engineer (Agentic Models)

Middle • 3+ years expAmsterdam • Remote (Germany, United Kingdom, Cyprus, Czech Republic, Netherlands) • Middle
Python
AI/ML
AI Agents
Dagster
Fine-tuning
Kubeflow
Langfuse
LLM
MLFlow
PyTorch
Tokenization
Megatron-LM
NVIDIA NeMo
Post-training
Pre-training
SFT
Function Calling
DevOps
Kubernetes
SLURM
Apply
Report

Databricks

databricks.com
Databricks is an American software company founded in 2013 by the original creators of Apache Spark, Delta Lake and MLflow at the University of California, Berkeley. Its data intelligence platform unifies data engineering, warehousing, governance, analytics and machine learning on a lakehouse architecture that stores open table formats in a customer's own cloud object storage. Headquartered in San Francisco and available on AWS, Azure and Google Cloud, the platform is used by tens of thousands of organisations to build pipelines, dashboards and production AI applications.
databricks.com • HQ: San Francisco, United States • MLOps • Machine Learning • Big Data • Data Engineering • Artificial Intelligence • Data & Analytics • Data Warehousing • 5000+ employees • Est. 2013
HQ: San Francisco, United States • MLOps • Machine Learning • Big Data • Data Engineering • Artificial Intelligence • Data & Analytics • Data Warehousing • 5000+ employees • Est. 2013
Verified live · 26 min ago 2 months ago

Staff Software Engineer, AI Runtime

$220k – $412k per year (Estimated) • Staff+ • 10+ years expIn office (Mountain View, United States) • Bachelor's Degree • Staff
Databricks
AI/ML
DeepSpeed
Fine-tuning
PyTorch
FSDP
InfiniBand
Megatron-LM
NVLink
Pre-training
Apply
Verified live · 26 min ago 2 months ago

Senior Software Engineer, AI Runtime

$183k – $355k per year (Estimated) • Senior • 5+ years expIn office (Mountain View, United States) • Bachelor's Degree • Senior
Databricks
AI/ML
DeepSpeed
Fine-tuning
PyTorch
FSDP
InfiniBand
Megatron-LM
NVLink
Pre-training
Apply
Verified live · 26 min ago 1 month ago

Sr. Engineering Manager, AI Runtime

$210k – $403k per year (Estimated) • Lead • 8+ years expIn office (Mountain View, United States) • Bachelor's Degree • Senior
Databricks
AI/ML
DeepSpeed
Fine-tuning
LLM
PyTorch
FSDP
Megatron-LM
NCCL
Apply
Report

Nimble Robotics

nimble.ai
Nimble (Nimble Robotics) is an AI-powered robotics and autonomous third-party logistics (3PL) company headquartered in San Francisco, California. Founded in 2017 by Simon Kalouche, the company develops fully autonomous fulfillment networks powered by intelligent picking, packing, and sorting robots ("Superhumanoids") alongside cloud supply chain management software.
nimble.ai • HQ: San Francisco, United States • Transportation & Logistics • Commerce • Logistics • 51-200 employees • Est. 2017
HQ: San Francisco, United States • Transportation & Logistics • Commerce • Logistics • 51-200 employees • Est. 2017
Top 25% payVerified live · 9 hours ago 13 days ago

Senior/Staff Software Engineer, Infrastructure (ML)

$210k – $300k per year • Staff+ • 4+ years expIn office (San Francisco, United States) • Bachelor's Degree • Staff
C++
Python
Rust
C++
PyTorch C++
AI/ML
CUDA Toolkit
JAX
PyTorch
DeepSpeed
FSDP
Megatron-LM
NCCL
DevOps
Kubernetes
Apply
Report

Hitachi

hitachi.com
Hitachi is a major Japanese multinational conglomerate specializing in industrial equipment, power systems, railway infrastructure, and digital technologies. Through its Social Innovation Business and Lumada platform, the company combines operational technology with enterprise IT to modernize critical infrastructure globally. Operating across diverse sectors - including energy, mobility, healthcare, and green technologies - Hitachi drives digital transformation and sustainable development worldwide.
hitachi.com • HQ: Tokyo, Japan • Hardware • Industrial Machinery • Manufacturing • 5000+ employees
HQ: Tokyo, Japan • Hardware • Industrial Machinery • Manufacturing • 5000+ employees
Senior • 10+ years expRemote/Hybrid (China) • Master's Degree • Full-Time • Senior
Apache Kafka
Databricks
AI/ML
RAG
AI Agents
Flink
TensorRT
Hugging Face
Megatron-LM
NVIDIA NeMo
Seldon Core
DevOps
Kubernetes
Splunk
Apply
Report

Figure AI

figure.ai
Figure AI is an American robotics company founded in 2022 that develops general-purpose humanoid robots for commercial and household work. Its Figure 02 and Figure 03 machines combine custom actuators, battery systems and hands with the in-house Helix vision-language-action model, which lets one neural network drive both perception and motor control from natural-language instructions. Headquartered in San Jose, California, the company runs the BotQ manufacturing facility for high-volume production and has piloted its robots on logistics and automotive assembly work with commercial partners.
figure.ai • HQ: San Jose, United States • Manufacturing • Humanoid Robots • Artificial Intelligence • Robotics • Robotics AI • 501-1000 employees • Est. 2022
HQ: San Jose, United States • Manufacturing • Humanoid Robots • Artificial Intelligence • Robotics • Robotics AI • 501-1000 employees • Est. 2022
Top 25% payVerified live · 7 hours ago 18 days ago

Helix AI Engineer, Training Performance

$200k – $400k per year • Middle • 3+ years expIn office (San Jose, United States) • Bachelor's Degree • Full-Time • Middle
C++
Python
C++
PyTorch C++
AI/ML
CUDA Toolkit
DeepSpeed
JAX
PyTorch
vLLM
AWS Trainium
FSDP
InfiniBand
Megatron-LM
NCCL
NVLink
TPU
AI Agents
Apply
Report

Nuancelabs

nuancelabs.ai
Face-to-face AI interaction that feels human. We are building a human foundation model with emotional intelligence — it reads tone, expression, and even hesitation, and responds in real time.
nuancelabs.ai • Seattle • Software • Conversational AI • Artificial Intelligence • 51-200 employees • Est. 2025
Seattle • Software • Conversational AI • Artificial Intelligence • 51-200 employees • Est. 2025
Top 25% payVerified live · 1 day ago 2 months ago

Member of Technical Staff — Pretraining Infra (Experienced)

$300k – $400k per year • Staff+ • In office (Seattle, United States) • Visa sponsorship • Staff
DeepSpeed
LLM
Multimodal AI
PyTorch
FSDP
Megatron-LM
NCCL
Synthetic Data
Post-training
DevOps
HPC
Apply
Report

Lightning AI

lightning.ai
Lightning AI is an artificial intelligence technology company headquartered in New York City, New York, and established in 2019. The organization provides a unified platform and the open-source PyTorch Lightning framework to facilitate the building, training, and deployment of machine learning models. It operates globally by offering scalable cloud infrastructure and development tools to individual researchers and enterprise teams across various industries.
lightning.ai • HQ: New York, United States • Cloud Management • Software • Cloud Computing • Information Technology • MLOps • Machine Learning • Artificial Intelligence • Est. 2019
HQ: New York, United States • Cloud Management • Software • Cloud Computing • Information Technology • MLOps • Machine Learning • Artificial Intelligence • Est. 2019
Verified live · 2 hours ago 19 days ago

Senior Research Engineer, LLM Training & Post-Training

$152k – $295k per year (Estimated) • Equity • Senior • Remote/Hybrid (New York, United States) • Master's Degree • Senior
Python
AI/ML
CUDA Toolkit
DeepSpeed
Fine-tuning
Lightning Fabric
LLM
PEFT
PyTorch
PyTorch Lightning
Reinforcement Learning
RLHF
SGLang
TensorRT
Transformers
TRL
vLLM
DPO
FSDP
GRPO
Hugging Face
Megatron-LM
Post-training
PPO
SFT
DevOps
Platform Engineering
Apply
Report

Applied Intuition

appliedintuition.com
Applied Intuition is a Silicon Valley software company founded in 2017 that develops digital infrastructure and simulation platforms for autonomous systems and physical AI. The platform provides advanced tools for simulation, synthetic data generation, and vehicle software management, allowing engineers to develop, test, and safely deploy self-driving technology. It serves leading global automotive manufacturers, defense organizations, and commercial transport companies across sectors like trucking, agriculture, and mining.
appliedintuition.com • HQ: Mountain View, United States • Smart Mobility • Manufacturing • Automotive Manufacturing • Artificial Intelligence • Autonomous Driving • 1001-5000 employees • Est. 2017
HQ: Mountain View, United States • Smart Mobility • Manufacturing • Automotive Manufacturing • Artificial Intelligence • Autonomous Driving • 1001-5000 employees • Est. 2017
Top 25% payVerified live · 2 hours ago 19 days ago

AI Performance Engineer

$215k – $285k per year • In office (Sunnyvale, United States) • Full-Time
C++
Python
C++
PyTorch C++
AI/ML
DeepSpeed
ONNX
Quantization
Ray
TensorRT
Triton Inference Server
Triton
FSDP
Megatron-LM
NCCL
CUDA Toolkit
OpenCV
PyTorch
CUDA
DevOps
Kubernetes
SLURM
Robotics
ROS
Apply
Top 25% payVerified live · 2 hours ago 19 days ago

Machine Learning Performance Engineer - Offboard Training & Inference

$215k – $285k per year • In office (Sunnyvale, United States) • Full-Time
C++
Python
C++
PyTorch C++
AI/ML
DeepSpeed
ONNX
Quantization
Ray
TensorRT
Triton Inference Server
FSDP
Megatron-LM
NCCL
CUDA Toolkit
OpenCV
PyTorch
DevOps
Kubernetes
SLURM
Robotics
ROS
Apply
Report

Eka Care

eka.care
Eka Care provides a personal health record and clinic management platform. Patients store prescriptions and reports while doctors manage consultations digitally. The company integrates with national digital health infrastructure.
eka.care • HQ: Bengaluru, India • Digital Health • Hospitals & Clinics • Health Care • Est. 2020
HQ: Bengaluru, India • Digital Health • Hospitals & Clinics • Health Care • Est. 2020
Verified live · 11 hours ago 21 day ago

ML Research Engineer - Pre-training (LLMs)

$21k – $62k per year (Estimated) • Junior • 2+ years expIn office (Bengaluru, India) • Full-Time • Junior
Megatron-LM
NVIDIA NeMo
Pre-training
Mixture of Experts
Apply
Report

Micron Technology

micron.com
Micron Technology is an American semiconductor company founded in Boise, Idaho in 1978 and one of only a handful of manufacturers of both DRAM and NAND flash memory. It designs and fabricates memory and storage products used in data centres, personal computers, smartphones, cars and industrial equipment, and sells them under the Micron and consumer-facing Crucial brands. High-bandwidth memory for AI accelerators has become the fastest-growing part of the business, and the company operates fabrication and assembly sites across the United States and Asia.
micron.com • HQ: Boise, United States • Data Centers • Semiconductor Manufacturing • Consumer Electronics • Hardware • Storage • Digital Storage • Semiconductors • 5000+ employees • Est. 1978
HQ: Boise, United States • Data Centers • Semiconductor Manufacturing • Consumer Electronics • Hardware • Storage • Digital Storage • Semiconductors • 5000+ employees • Est. 1978
Verified live · 1 hour ago 27 days ago

MTS, AI Engineering, SMAI

Senior • 5+ years expIn office (Taichung, Taoyuan, Taiwan) • PhD • Full-Time • Senior
C++
Java
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
AI Agents
AutoGen
Chain-of-Thought
Computer Vision
CUDA
CUDA Toolkit
DeepSpeed
Fine-tuning
Function Calling
Kubeflow
LangChain
LangGraph
LlamaIndex
LLM
LoRA
OpenCL
PEFT
Prompt Engineering
PyTorch
QLoRA
Ray
RLHF
Scikit-learn
TensorFlow
TensorRT
TensorRT-LLM
Triton
vLLM
Transformers
FSDP
Megatron-LM
NVLink
SFT
Frontend
Sass
Mobile
Metal
DevOps
CI/CD
Docker
Git
Jenkins
Kubernetes
SLURM
HPC
Apply
Verified live · 1 hour ago 1 month ago

Member of Technical Staff (MTS), Machine Learning, SMAI

$136k – $298k per year (Estimated) • Staff+ • 10+ years expIn office (Singapore) • Full-Time • Staff
C++
Java
Python
C++
PyTorch C++
TensorFlow C++
Databases
Snowflake
AI/ML
AI Agents
AutoGen
Chain-of-Thought
Computer Vision
CrewAI
CUDA
CUDA Toolkit
DeepSpeed
Fine-tuning
Function Calling
Kubeflow
LangChain
LangGraph
LlamaIndex
LLM
LoRA
PEFT
Prompt Engineering
PyTorch
QLoRA
Ray
RLHF
Scikit-learn
TensorFlow
TensorRT
TensorRT-LLM
Triton
vLLM
Transformers
FSDP
Megatron-LM
NVLink
SFT
DevOps
CI/CD
Docker
GCP
Git
Jenkins
Kubernetes
SLURM
HPC
Analytics
ETL/ELT
Apply
Verified live · 1 hour ago 2 months ago

Member of Technical Staff, AI Engineering

$162k – $328k per year (Estimated) • Staff+ • 10+ years expIn office (Boise, United States) • Bachelor's Degree • Full-Time • Staff
C++
Java
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
AI Agents
AutoGen
Chain-of-Thought
CUDA
CUDA Toolkit
DeepSpeed
Fine-tuning
Function Calling
LangChain
LangGraph
LlamaIndex
LLM
LoRA
OpenCL
PEFT
Prompt Engineering
PyTorch
QLoRA
RLHF
Scikit-learn
TensorFlow
TensorRT
TensorRT-LLM
vLLM
Transformers
Edge AI
FSDP
Megatron-LM
NVLink
SFT
Computer Vision
Kubeflow
Ray
Triton
Frontend
Sass
DevOps
CI/CD
Docker
Git
Jenkins
Kubernetes
SLURM
HPC
Apply
Report

Thinking Machines Lab

thinkingmachines.ai
Thinking Machines Lab is an artificial intelligence research and product company based in San Francisco and founded in 2025. The company develops multimodal AI systems and open-weights models, such as Inkling, alongside developer tools like Tinker for model fine-tuning. It operates as a public benefit corporation focused on human-AI collaboration and open science, supported by significant venture capital investment.
thinkingmachines.ai • HQ: San Francisco, United States • MLOps • Machine Learning • Multimodal AI • LLM & Generative AI • Artificial Intelligence • 201-500 employees • Est. 2025
HQ: San Francisco, United States • MLOps • Machine Learning • Multimodal AI • LLM & Generative AI • Artificial Intelligence • 201-500 employees • Est. 2025
Top 25% payVerified live · 1 hour ago 27 days ago

Research Engineer, Infrastructure, Numerics

$350k – $475k per year • Remote/Hybrid (San Francisco, United States) • Relocation • Visa sponsorship • Full-Time
JAX
PyTorch
DeepSpeed
Megatron-LM
Apply
Top 25% payVerified live · 1 hour ago 27 days ago

Research Engineer, Infrastructure, Training Systems

$350k – $475k per year • Remote/Hybrid (San Francisco, United States) • Relocation • Visa sponsorship • Full-Time
JAX
PyTorch
DeepSpeed
Megatron-LM
Apply
Report

Meesho

meesho.com
Meesho is an Indian e-commerce company headquartered in Bengaluru and founded in 2015. The company operates an online marketplace that enables small businesses and individual entrepreneurs to sell products such as apparel, home goods, and electronics directly to consumers. Originally established as a social commerce platform facilitating sales through networks like WhatsApp and Facebook, it has evolved into a large-scale retail ecosystem serving millions of users across India.
meesho.com • HQ: Bengaluru, India • B2B Commerce • Apparel & Fashion • Consumer Goods • Commerce • Marketplaces • Social Commerce • Est. 2015
HQ: Bengaluru, India • B2B Commerce • Apparel & Fashion • Consumer Goods • Commerce • Marketplaces • Social Commerce • Est. 2015
Verified live · 7 hours ago 1 month ago

Engineering Manager – AI Engineering

$42k – $93k per year (Estimated) • Lead • 9+ years expIn office (Bengaluru, India) • Bachelor's Degree • Full-Time • Staff
C++
Python
Rust
C++
PyTorch C++
AI/ML
CUDA
CUDA Toolkit
DeepSpeed
Flink
Knowledge Distillation
LLM
PyTorch
Quantization
Ray
SGLang
Spark
TensorRT
TensorRT-LLM
vLLM
AI Agents
FSDP
LLMOps
Megatron-LM
DevOps
FinOps
Google GKE
Kubernetes
GCP
Apply
Report

Together AI

together.ai
Together AI (Together Computer, Inc.) is a full-stack AI infrastructure and cloud platform headquartered in San Francisco, California. Founded in 2022 by prominent AI researchers and system engineers - including CEO Vipul Ved Prakash, CTO Ce Zhang, Chief Scientist Tri Dao (co-creator of FlashAttention), Chris Ré, and Percy Liang - the company operates as an "AI Native Cloud" designed to train, fine-tune, and deploy open-source generative AI models at scale with high performance and optimized unit economics.
together.ai • HQ: San Francisco, United States • Hardware • Events & Ticketing • Corporate Events • AI Infrastructure • Artificial Intelligence • 501-1000 employees • Est. 2022
HQ: San Francisco, United States • Hardware • Events & Ticketing • Corporate Events • AI Infrastructure • Artificial Intelligence • 501-1000 employees • Est. 2022
Top 25% payVerified live · 11 hours ago 1 month ago

Research Engineer, Large-Scale Training

$200k – $290k per year • In office (San Francisco, United States) • Full-Time
Python
AI/ML
Fine-tuning
PyTorch
Together AI
CUDA
CUDA Toolkit
DeepSpeed
Mamba
Triton
FSDP
Megatron-LM
NCCL
Apply
Report

ConveGenius

convegenius.com
ConveGenius is an education management company that provides services related to sustainable education for K-10 categories through impact-driven programs, an adaptive learning platform, and data dashboards.
convegenius.com • HQ: Noida, India • Primary Education • Est. 2013
HQ: Noida, India • Primary Education • Est. 2013
1 month ago

ML Engineer

$29k – $121k per year (Estimated)In office (Noida, India)
DeepSpeed
Fine-tuning
LoRA
PEFT
PyTorch
QLoRA
RLHF
Transformers
TRL
DPO
Hugging Face
Megatron-LM
SFT
Apply
$28k – $118k per year (Estimated)In office (Noida, India)
DeepSpeed
Fine-tuning
LoRA
PEFT
Perplexity
PyTorch
QLoRA
RLHF
Transformers
TRL
DPO
FSDP
Hugging Face
Megatron-LM
PPO
SFT
Apply
1 month ago

AI Platform Engineer

$26k – $109k per year (Estimated)In office (Noida, India)
Databricks
AI/ML
AWS Bedrock
DeepSpeed
Fine-tuning
LangSmith
LLM
LoRA
NLP
PyTorch
QLoRA
RAG
RLHF
vLLM
PEFT
Amazon SageMaker
DPO
Megatron-LM
AI Agents
Function Calling
TGI
DevOps
Amazon EKS
AWS
Azure
CI/CD
Docker
GCP
Google GKE
Kubernetes
OpenTelemetry
Vector
Apply
Report
Megatron-LM Jobs - Remote & On-site
Frequently asked questions
How many Megatron-LM jobs are available on Alion?
There are currently 75 active listings from 43 hiring companies on Alion, refreshed daily.
What salary do Megatron-LM jobs pay?
The average salary is around $329,084 per year based on the current listings.
Are remote options available for Megatron-LM jobs?
Yes, many of these roles offer remote or hybrid work, and you can filter listings by remote, relocation and visa sponsorship.
How do I apply for Megatron-LM jobs?
Open any listing on Alion and apply directly with the employer through the provided link - no lengthy sign-up required.