368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$200k – $350k per year
Location
Remote (United States)
Seniority
Staff · 10+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Positron AI builds purpose-built inference servers for large language models, aiming for much higher energy efficiency than general-purpose GPU systems. Its Atlas hardware uses field-programmable logic and a memory-centric design to serve transformer models at high token throughput. The company was founded in 2023 and assembles its systems in the United States.

About Positron AI

Positron AI is building next-generation AI inference accelerators designed from the ground up for low-latency, high-throughput large language model inference. Our first-generation ASIC, Asimov, is a cutting-edge accelerator targeting frontier AI workloads, with additional generations already underway.

Role Overview

Positron AI is looking for a Technical Product Manager to own AI inference and software technical product planning end to end. In this role, you will be the person who translates where models and inference systems are heading into concrete, well-scoped requirements for our inference software stack, spanning model coverage, numerics, inference-engine features and modes, serving-stack capabilities, and our managed service.

This is a deeply technical planning role that sits at the intersection of engineering, go-to-market, and the broader inference ecosystem. You will track the model frontier as a discipline, convert that movement into engineering requests before it becomes a customer escalation, and serve as the connective tissue between our engineering organization, our GTM teams, and our ecosystem partners. You will also be expected to use agentic AI daily as a core part of how the planning function operates.

Key Responsibilities

  • Be the leader in Positron for all aspects of AI Inference & Software Technical Product Planning.
  • Write requirements for our inference software stack spanning model coverage, numerics, inference-engine features/modes, serving-stack features/modes, and managed-service capabilities.
  • Keep our planning ahead of where models and inference systems are going - track the model frontier as a discipline and convert movement into requirements before it becomes a customer escalation.
  • Work with engineering and GTM on communicating a roadmap.
  • Work closely with GTM teams to understand customer and market needs.
  • Define and manage the software product lifecycle: versions and release trains, feature modes, model catalog and deprecation policy.
  • Creative competitive briefings.
  • Work with our key ecosystem partners - model labs, open-source runtimes, serving and orchestration partners - to understand their technology roadmaps.
  • Create scope/feasibility frameworks to convert model and inference-system innovations into tangible engineering requests.
  • Ensure our software products are well documented.
  • Streamline & automate the product planning processes using agentic AI.

Required Qualifications

  • 10+ years experience with a strong mix of these types of technical skills:
    • ML systems, inference infrastructure, or serving-stack engineering with direct ownership of performance or architecture trade-offs - transformer internals at the operator level: attention variants, MoE routing, KV-cache mechanics, quantization formats and their hardware implications.
    • Production inference serving at scale touching on multi-tenancy, latency SLAs (TTFT/TPOT), batching and scheduling, disaggregated serving, KV-cache management, and observability.
    • The open-source inference ecosystem - runtimes (vLLM/SGLang-class), kernels, model ingestion, and how models are released, quantized, and adopted in practice.
    • Performance analysis spanning models & systems: utilization reasoning, tokens per $ and per W arithmetic, benchmark design - able to build and defend the math personally.
    • Competitive landscaping and analysis of inference providers and serving stacks, at a model lab, inference API provider, or AI hardware company.
  • 10+ years with a strong mix of these type of personal & analytic skills:
    • Quick learner with the ability to span model-architecture details up to fleet-scale serving systems. Stays current with the model and inference landscape.
    • Excellent communication and people skills, comfortable navigating uncertainty, and driving a process of idea & decision socialization.
    • Comfortable being the most technically grounded person in a GTM room and the most market-aware person in an engineering room.
    • Owning decision history: be the personification of documented answers to “why did we choose X”, and do so with our executive leadership.
    • Demonstrated daily, hands-on use of agentic AI in real technical work - building and running agent workflows for research, analysis, and drafting req’s - with the judgment to verify and own everything the agents produce.

Leveling & Scope

While this role is currently posted at a specific level, we are a growth-oriented organization and are open to hiring at a more senior level for the right candidate. Please note that this job description serves as a focused but generalized overview of the role; specific responsibilities and impact expectations will be tailored to the experience and seniority of the final hire.

Why Join Us?

  • You will shape the software roadmap for a purpose-built inference accelerator, working at the layer where model architecture, systems performance, and real customer workloads meet.
  • You will have unusually direct influence, defining what gets built across the inference stack and seeing it land in silicon-backed products that compete on performance per dollar and performance per watt.

Compensation & Benefits

The base salary range for this role is $200,000 - $350,000.

Please note that the figures provided represent the base salary range only and do not include other elements of our total compensation package, equity, or comprehensive benefits.

At Positron AI, we value the unique expertise each candidate brings. While the range above reflects our typical expectation for the position, we reserve the flexibility to exceed this range for candidates whose specialized skills, significant experience, or unique qualifications fall outside the standard scope of the role. Final offers are determined based on a variety of factors, including internal equity, and individual impact.

Benefits & Perks

We want you to do your best work and feel confident that you and your family are taken care of. That means comprehensive coverage, real time to rest, and support for your future.

Health and wellness

  • Fully company-paid medical, dental, and vision insurance for you and your dependents
  • Company-paid life and disability coverage, with voluntary options to add more
  • Supplemental hospital, critical illness, and accident coverage available

Time off and flexibility

  • Unlimited paid time off, we encourage everyone to truly unplug and recharge
  • 13 paid company holidays
  • Remote-first culture with a company-provided computer and home office setup

Compensation and future

  • Competitive salary and equity
  • 401(k) with company matching, eligible from day one

Visa Support

This position is open to candidates currently authorized to work in the U.S. We cannot provide new visa sponsorship for this role but are open to facilitating H-1B visa transfers for eligible candidates.

Equal Opportunity Employer. If you're excited about the role but don't meet every bullet, we'd still love to hear from you.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$34k – $82k per year (Estimated) • Remote/Hybrid • Full-Time • 10+ years exp • Bengaluru • Hyderabad
Python
AI/ML
AI Agents
Anomaly Detection
LangChain
LlamaIndex
LLM
LLM Guardrails
RAG
DevOps
AWS
Azure
CI/CD
GCP
Platform Engineering
Vector
Management
ServiceNow
Apply
$28k – $116k per year (Estimated) • Remote/Hybrid • Full-Time • Bengaluru • Hyderabad
Python
SQL
Python
FastAPI
Pydantic
Databases
FAISS
OpenSearch
Pinecone
SAP HANA
Snowflake
AI/ML
A2A
AI Agents
AWS Bedrock
AWS Bedrock AgentCore
Claude
Claude Code
Cline
Copilot
Cursor
DeepEval
Embeddings
Gemini
Hallucination
LangChain
LangGraph
Llama
LLM
LLM Guardrails
Model Context Protocol
OCR
OpenAI
Promptfoo
RAG
DevOps
AWS
Azure
GCP
GitHub
Management
Power Automate
Marketing
Salesforce
Apply
In office • Full-Time • Sydney
Python
SQL
Databases
Databricks
Snowflake
AI/ML
AI Agents
Copilot
DevOps
AWS
Azure
GCP
GitHub
Apply
$115k – $228k per year (Estimated) • Remote/Hybrid • 6+ years exp • Bachelor's Degree • Scottsdale
MATLAB
SystemVerilog
TCL Scripting
Verilog
VHDL
MATLAB
Simulink
AI/ML
Claude
Claude Code
Copilot
Cursor
Quantization
DevOps
Git
GitHub
Chips/EDA
HDL Coder
Xilinx Vivado
Apply
$110k – $219k per year (Estimated) • Equity • Remote/Hybrid • 6+ years exp • Bachelor's Degree • Boston
MATLAB
SystemVerilog
TCL Scripting
Verilog
VHDL
MATLAB
Simulink
AI/ML
Claude
Claude Code
Copilot
Cursor
Quantization
DevOps
Git
GitHub
Chips/EDA
HDL Coder
Xilinx Vivado
Apply
$200k – $350k per year • In office • Full-Time • 12+ years exp • Austin
SystemVerilog
DevOps
CI/CD
HPC
Chips/EDA
Cadence Palladium
Formal Verification
Siemens Veloce
Synopsys ZeBu
UVM
Apply
$200k – $300k per year • Remote • Full-Time • 12+ years exp
Cybersecurity
Threat Modeling
Apply
$150k – $300k per year • Remote • Full-Time • 12+ years exp
Chips/EDA
Ansys RedHawk
Synopsys PrimeTime
Apply
$200k – $300k per year • Remote • Full-Time • 12+ years exp
AI/ML
LLM
DevOps
Vector
Chips/EDA
Synopsys PrimePower
Apply
$190k – $280k per year • Remote • Full-Time • 12+ years exp
Python
Chips/EDA
Cadence Genus
Synopsys Design Compiler
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.