368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$152k – $274k per year (Estimated)
Location
Remote/Hybrid (New York, United States)
Seniority
Senior · 7+ years exp
Overview
Company
Impact
Profile match
Lightning AI is an artificial intelligence technology company headquartered in New York City, New York, and established in 2019. The organization provides a unified platform and the open-source PyTorch Lightning framework to facilitate the building, training, and deployment of machine learning models. It operates globally by offering scalable cloud infrastructure and development tools to individual researchers and enterprise teams across various industries.

Who We Are

Lightning AI is the company behind PyTorch Lightning. Founded in 2019, we build an end-to-end platform for developing, training, and deploying AI systems-designed to take ideas from research to production with less friction.

Through our merger with Voltage Park, a neocloud and AI Factory, Lightning AI combines developer-first software with cost-efficient, large-scale compute. Teams get the tools they need for experimentation, training, and production inference, with security, observability, and control built in.

We serve solo researchers, startups, and large enterprises. Lightning AI operates globally with offices in New York City, San Francisco, Seattle, and London, and is backed by Coatue, Index Ventures, Bain Capital Ventures, and Firstminute.

The Way We Work

The people who thrive here are builders who move fast, communicate openly, take ownership, and continuously improve themselves, their teams, and our company. Here's what that looks like in practice:

  • Move with Urgency: We move quickly, make thoughtful decisions, and keep momentum. We value action over perfection and learn by shipping.
  • Take Ownership: We own outcomes, not just our individual work. We make decisions that move the company forward and follow through.
  • Communicate Openly: We communicate directly, seek to understand, and create clarity for others. Honest conversations help us move faster together.
  • Build Great Teams: We lead by example, empower others, and create healthy teams where people can do their best work.
  • Raise the Bar: We're always improving ourselves. We learn from feedback, consistently challenge ourselves to grow, and focus on the work that matters most.
  • Think Long-Term: We design for what's next. We create scalable systems, simplify complexity, and use AI and automation to amplify our impact.

What We're Looking For

We’re looking for a Senior Product Manager to own Lightning AI’s experimentation and post-training product end to end-from product strategy and roadmap through launch, adoption, pricing, and go-to-market.

This is a role focused on how AI researchers and engineers turn an idea into a high-quality, validated model. You’ll define the workflow for running and comparing experiments, managing fine-tuning and reinforcement-learning workloads, evaluating model quality, understanding failures, and selecting the right model or checkpoint for production.

You’ll work at the intersection of developer tooling, AI infrastructure, and model quality. The right candidate understands how modern AI teams work today: notebooks, training jobs, experiment trackers, checkpoints, evaluation suites, spreadsheets, and custom internal tooling. You can identify where those workflows break down and turn them into a cohesive, minimal product experience.

You should be able to move fluidly between designing an intuitive developer workflow, discussing distributed training and artifact lineage with engineers, evaluating model outputs, and explaining the product’s value to customers and sales teams.

This role requires unusually high ownership. You will work directly with engineering, customers, and the executive team; develop strong product opinions; create your own product artifacts; and drive work forward without waiting for another function to define the next step.

You will join the Product Team, report to our VP of Product, and work directly with our executive team as we grow this business.

This is a hybrid role based in our New York City or San Francisco office, with an in-office expectation of two days per week.

What You’ll Do

  • Own the product vision and roadmap for post-training and experimentation - what we build, what we integrate with, what we don't build, and in what order
  • Understand how ML engineers and AI researchers actually work today: the jobs they run, the comparisons they make, the failures they debug, and the handoffs that break down between research and production - then build the product that makes that workflow coherent
  • Develop a strong point of view on where Lightning should build differentiated experiences versus integrate with the existing ecosystem of experiment trackers, evaluation frameworks, data tools, and model registries
  • Work directly with engineers from problem definition through architecture, implementation, and launch - understand the constraints, help shape the solutions, don't hand off requirements and wait
  • Use the product yourself; inspect failed workflows, read logs, identify friction, and remove it without waiting for someone to surface it
  • Own model evaluation as a product function - write evals, assess outputs, and let quality signals drive roadmap decisions
  • Design pricing and packaging in partnership with Growth and Finance - model unit economics, run experiments, and make calls that affect both adoption and margin
  • Build workflows that help teams collaborate: share results, compare models, move work from research into production, and maintain enough lineage that decisions can be explained and reproduced
  • Be the product voice in GTM - sales positioning, technical objection handling, and developer-facing content that builds credibility with ML engineers and platform teams
  • Define and instrument the metrics that matter across activation, iteration speed, compute consumption, retention, and expansion

What You’ll Need

  • 7+ years of product management experience, including at least 3 years building infrastructure, platform, developer-tooling, or machine-learning products.
  • Hands-on experience building products for ML engineers, AI researchers, or data scientists.
  • A detailed understanding of experimentation and post-training workflows, including training jobs, checkpoints, metrics, artifacts, experiment comparison, reproducibility, and model evaluation.
  • Experience with one or more modern post-training techniques, such as supervised fine-tuning, preference optimization, reinforcement learning, distributed training, or hyperparameter optimization.
  • Experience designing or working closely with model evaluations. You understand that model quality is multidimensional and know how qualitative and quantitative signals should inform product decisions.
  • Strong product and interaction judgment. You can simplify technically complex workflows, develop a clear information architecture, and create usable product experiences without depending heavily on a dedicated design function.
  • Enough technical depth to work directly with infrastructure and ML engineers on APIs, SDKs, execution systems, distributed workloads, observability, data and artifact management, and failure handling.
  • A record of end-to-end ownership. You do not stop at roadmap definition; you investigate problems personally, create the necessary artifacts, drive decisions, support implementation, validate the result, and help bring the product to market.
  • Strong prioritization and willingness to say no. You can identify the narrowest valuable product wedge and resist building a broad collection of loosely connected features.
  • Experience owning pricing, packaging, or consumption-based products, including working with unit economics and making decisions that affect adoption and margin.
  • Strong written and verbal communication. You are equally comfortable writing a technical product specification, reviewing a developer workflow, presenting to executives, and explaining the product to a customer.
  • High bias for action and comfort operating in ambiguous, fast-moving environments with limited process and support.
  • BS in Computer Science, Engineering, or equivalent practical experience.

Bonus Points

  • Experience at an AI infrastructure company, neocloud, hyperscaler ML platform, experiment-tracking company, or developer-tooling startup.
  • Familiarity with PyTorch, PyTorch Lightning, distributed training, GPU infrastructure, or large-scale fine-tuning.
  • Experience with tools such as Weights & Biases, MLflow, Hugging Face, Ray, Slurm, Kubernetes, or comparable internal ML platforms.
  • Experience building collaborative workflows for teams moving models from research through evaluation and production.

Compensation

We are committed to offering competitive compensation that reflects the value each team member brings to our mission. Final offers are based on factors such as experience, skills, geographic location, and role expectations. In addition to base salary, our total rewards package for eligible roles includes a discretionary bonus, a meaningful equity component, and comprehensive benefits.

The anticipated annual base salary range for this role is:

$160,000—$275,000 USD

Benefits and Perks

We offer a comprehensive and competitive benefits package designed to support our employees’ health, well-being, and long-term success:

  • Comprehensive Health Coverage: Medical, dental, and vision coverage for employees and eligible dependents.
  • Meaningful Equity: RSUs that give employees a stake in the company's long-term success.
  • Retirement Savings: 401(k) matching (U.S.) and pension contributions (U.K.).
  • Flexible Time Off: Unlimited PTO, company holidays, and floating holidays to support work-life balance.
  • Company-Wide Winter Break: Two weeks of company closure each winter to disconnect and recharge.
  • Paid Parental & Family Leave: Paid leave to support you and your family through life's important moments.
  • Professional Development: Annual learning and development allowance to support your professional growth.
  • Wellness Benefits: Wellness and work-from-home stipends to support your physical and mental well-being.
  • Sabbatical Program: Four weeks of paid sabbatical leave after four years of service.
  • Flexible Work: Flexible schedules and a hybrid work model for our office-based teams.
  • In-Office Meals: Complimentary meals at our office hubs.

Benefits may vary by location, team, and role.

At Lightning AI, we are committed to fostering an inclusive and diverse workplace. We believe that diverse teams drive innovation and create better products. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other protected characteristic. We are dedicated to building a culture where everyone can thrive and contribute to their fullest potential.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
New York
$120k – $250k per year • Equity 0.2–1% • In office • Full-Time • Master's Degree • Seattle
Python
AI/ML
Diffusion Models
Fine-tuning
PyTorch
Self-Supervised Learning
Apply
$95k – $240k per year (Estimated) • In office • Full-Time • Singapore
Python
SQL
AI/ML
Computer Vision
PyTorch
Scikit-learn
TensorFlow
AI Agents
Analytics
Power BI
Tableau
Apply
$140k – $225k per year • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Seattle
TypeScript
Python
Python
FastAPI
Pydantic
Databases
pgvector
PostgreSQL
Redis
AI/ML
Fine-tuning
LangGraph
LLM
RAG
LangChain
AutoGen
CrewAI
Knowledge Distillation
NLP
Reranking
Human-in-the-Loop
LLM Evaluation
LLM Guardrails
OCR
Structured Outputs
AI Agents
Function Calling
Speech Recognition
DevOps
Docker
Terraform
AWS
Azure
Bicep
Robotics
Digital Twin
Marketing
Salesforce
Apply
$300k – $475k per year • Remote/Hybrid • Full-Time • 4+ years exp • San Francisco
Python
Rust
TypeScript
JavaScript
AI/ML
Fine-tuning
Frontend
React.js
Apply
ML Engineer 1 day ago
Remote/Hybrid • 3+ years exp
Python
SQL
AI/ML
AI Agents
Computer Vision
MLFlow
PyTorch
RAG
A2A
Model Context Protocol
DevOps
AWS
Azure
CI/CD
Docker
Docker Compose
GCP
Kubernetes
Analytics
ETL/ELT
Apply
Frontend Engineer 4 days ago
Equity • Remote/Hybrid • London
JavaScript
TypeScript
AI/ML
PyTorch
PyTorch Lightning
Frontend
React.js
Redux
DevOps
CI/CD
Apply
$104k – $217k per year (Estimated) • Equity • Remote/Hybrid • London
Python
AI/ML
CUDA Toolkit
Kubeflow
PyTorch
PyTorch Lightning
Ray
CUDA
InfiniBand
NCCL
DevOps
Grafana
Kubernetes
OpenTelemetry
Platform Engineering
Prometheus
SLURM
Apply
$105k – $221k per year (Estimated) • Equity • Remote/Hybrid • Bachelor's Degree • London
AI/ML
CUDA Toolkit
PyTorch
PyTorch Lightning
SGLang
vLLM
CUDA
DeepSpeed
Triton
FSDP
Hugging Face
Post-training
Apply
$140k – $190k per year • Equity • Remote/Hybrid • Contractor • 8+ years exp • San Francisco
AI/ML
PyTorch
PyTorch Lightning
Marketing
HubSpot
Salesforce
Apply
$149k – $269k per year (Estimated) • Equity • Remote/Hybrid • New York
Go
Python
AI/ML
PyTorch
PyTorch Lightning
DevOps
ArgoCD
CI/CD
Crossplane
GitOps
gRPC
Helm
Kubernetes
Kustomize
Platform Engineering
SLURM
Terraform
HPC
Apply
$220k – $350k per year • Remote/Hybrid • Full-Time • 15+ years exp • New York • Princeton
AI/ML
AI Agents
LLM Guardrails
Model Context Protocol
DevOps
Azure
Azure DevOps
CI/CD
GitHub
Platform Engineering
Design
Figma
Management
Jira
QA
Playwright
Apply
$80k – $115k per year • In office • Full-Time • PhD • New York
Apply
$60k – $116k per year (Estimated) • Remote/Hybrid • Bachelor's Degree • New York
Apply
UX Research Lead 1 hour ago
$192k – $288k per year • Remote/Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • New York
AI/ML
Hallucination
Design
Axure RP
Figma
Sketch
InVision
Apply
$120k – $240k per year (Estimated) • Equity • Remote/Hybrid • 5+ years exp • New York
C#
C++
Go
JavaScript
TypeScript
Frontend
React.js
Redux
DevOps
AWS
Azure
GCP
IAM
Cybersecurity
FedRAMP
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.