415,275open jobs
14,693companies
74,504added this week
Browse all
Salary
$200k – $300k per year
Location
Remote (United States)
Seniority
Architect · 15+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Positron AI builds purpose-built inference servers for large language models, aiming for much higher energy efficiency than general-purpose GPU systems. Its Atlas hardware uses field-programmable logic and a memory-centric design to serve transformer models at high token throughput. The company was founded in 2023 and assembles its systems in the United States.

About Positron AI

Positron AI is building next-generation AI inference accelerators designed from the ground up for low-latency, high-throughput large language model inference. Our first-generation ASIC, Asimov, is a cutting-edge accelerator targeting frontier AI workloads, with additional generations already underway.

Role Overview

We're looking for an exceptional Principal ASIC Architect to help define the future of AI inference hardware. You will collaborate across architecture, software, machine learning, systems, and silicon implementation to design next-generation AI accelerators capable of serving rapidly evolving frontier models. This role requires exceptional technical breadth, balancing mathematical understanding, ML trends, systems architecture, and practical silicon implementation. This may be the single most important technical hire after the Chief Architect.

Key Responsibilities

Architectural Exploration

  • Own architectural exploration across the evolving landscape of transformer models, MoE, sparse models, and long-context inference.
  • Evaluate emerging techniques including speculative decoding, continuous batching, KV-cache evolution, sub-quadratic and linear attention, and state-space models.
  • Assess memory architectures, communication architectures, and interconnects for next-generation accelerators.

Performance Modeling

  • Develop analytical models and build simulation frameworks to evaluate architectural tradeoffs.
  • Estimate latency, throughput, bandwidth, memory and compute utilization, scaling efficiency, power, area, and cost.

Hardware/Software Co-Design

  • Partner closely with compiler, runtime, system software, firmware, ASIC engineering, and ML research teams to align hardware direction with real workload needs.

ASIC Specification & Roadmap

  • Develop architecture documents and define subsystem interfaces.
  • Evaluate tradeoffs across competing architectural directions and guide the future product roadmap.

Industry Awareness

  • Track new LLMs, reasoning models, research papers, hardware competitors, academic research, and emerging algorithms to keep Positron's architecture ahead of the field.

AI-Enabled Engineering

  • Use AI to accelerate research, performance analysis, architecture exploration, simulation generation, and documentation.

Required Qualifications

  • 15+ years of experience in ASIC architecture.
  • Deep expertise in AI accelerator design and high-performance SoCs.
  • Hands-on experience building performance models for memory systems and interconnects.
  • Strong hardware/software co-design ability.
  • Excellent written and verbal communication skills.

Preferred Qualifications

  • Experience with GPU architecture.
  • A solid grasp of ML systems, distributed inference, and large-scale serving.
  • Deep knowledge of transformer internals.
  • Compiler experience.
  • A track record of research publications.

Leveling & Scope

While this role is currently posted at a specific level, we are a growth-oriented organization and are open to hiring at a more senior level for the right candidate. Please note that this job description serves as a focused but generalized overview of the role; specific responsibilities and impact expectations will be tailored to the experience and seniority of the final hire.

What Success Looks Like

Success in this role looks like helping define future Positron architectures and creating a scalable architectural direction that bridges software and hardware, influencing products that remain competitive for years despite rapidly evolving AI models.

Why Join Us?

  • Shape the core architecture of Positron's AI inference accelerators at a foundational level, with outsized influence over the company's technical direction for years to come.
  • Work at the intersection of frontier AI research and silicon implementation, tackling problems few architects in the industry get to own end to end.

Compensation & Benefits

The base salary range for this role is $200,000 - $300,000.

Please note that the figures provided represent the base salary range only and do not include other elements of our total compensation package, equity, or comprehensive benefits.

At Positron AI, we value the unique expertise each candidate brings. While the range above reflects our typical expectation for the position, we reserve the flexibility to exceed this range for candidates whose specialized skills, significant experience, or unique qualifications fall outside the standard scope of the role. Final offers are determined based on a variety of factors, including internal equity, and individual impact.

Benefits & Perks

We want you to do your best work and feel confident that you and your family are taken care of. That means comprehensive coverage, real time to rest, and support for your future.

Health and wellness

  • Fully company-paid medical, dental, and vision insurance for you and your dependents
  • Company-paid life and disability coverage, with voluntary options to add more
  • Supplemental hospital, critical illness, and accident coverage available

Time off and flexibility

  • Unlimited paid time off, so you can truly unplug and recharge
  • 13 paid company holidays
  • Remote-first culture with a company-provided computer and home office setup

Compensation and future

  • Competitive salary and equity
  • 401(k) with company matching, eligible from day one

Visa Support

This position is open to candidates currently authorized to work in the U.S. We cannot provide new visa sponsorship for this role but are open to facilitating H-1B visa transfers for eligible candidates.

Equal Opportunity Employer. If you're excited about the role but don't meet every bullet, we'd still love to hear from you.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
415,275 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$21k – $39k per year (Estimated) • In office • Full-Time • Moscow
SQL
Databases
FAISS
Milvus
Qdrant
AI/ML
A2A
Axolotl
Embeddings
Fine-tuning
Function Calling
GPTQ
Haystack
Hybrid Search
Knowledge Graph
Kubeflow
KV Cache
LangGraph
Llama
LLaMA-Factory
LLM
LoRA
Mistral
MLFlow
Model Context Protocol
QLoRA
Quantization
Qwen
RAG
Reranking
Speculative Decoding
TensorRT
TensorRT-LLM
vLLM
LangChain
PEFT
DevOps
Docker
Kubernetes
Vector
Apply
$110k – $263k per year (Estimated) • Remote • Full-Time
Python
AI/ML
CUDA
CUDA Toolkit
Edge AI
ONNX
PyTorch
Quantization
ROCm
TensorRT
Triton
vLLM
KV Cache
ONNX Runtime
Speculative Decoding
Apply
$67k – $129k per year (Estimated) • Remote • Full-Time
Python
AI/ML
CUDA
CUDA Toolkit
Edge AI
ONNX
PyTorch
Quantization
ROCm
TensorRT
Triton
vLLM
KV Cache
ONNX Runtime
Speculative Decoding
Apply
$80k – $206k per year (Estimated) • Remote • Full-Time
Python
AI/ML
CUDA
CUDA Toolkit
Edge AI
ONNX
PyTorch
Quantization
ROCm
TensorRT
Triton
vLLM
KV Cache
ONNX Runtime
Speculative Decoding
Apply
$70k – $131k per year (Estimated) • Remote • Full-Time
Python
AI/ML
CUDA
CUDA Toolkit
Edge AI
ONNX
PyTorch
Quantization
ROCm
TensorRT
Triton
vLLM
KV Cache
ONNX Runtime
Speculative Decoding
Apply
$200k – $350k per year • Remote • Full-Time • 10+ years exp
AI/ML
Quantization
SGLang
TPOT
vLLM
AI Agents
Mixture of Experts
KV Cache
Apply
$200k – $350k per year • In office • Full-Time • 12+ years exp • Austin
SystemVerilog
DevOps
CI/CD
HPC
Chips/EDA
Cadence Palladium
Formal Verification
Siemens Veloce
Synopsys ZeBu
UVM
Apply
ASIC Design Engineer 1 month ago
$225k – $350k per year • In office • Full-Time • 8+ years exp • Austin
Python
SystemVerilog
DevOps
Vector
Chips/EDA
Cocotb
Formal Verification
UVM
Apply
$200k – $350k per year • In office • Full-Time • 12+ years exp • Austin
SystemVerilog
Verilog
DevOps
HPC
Apply
$200k – $350k per year • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Austin
Perl
Python
SystemVerilog
Chips/EDA
Formal Verification
UVM
Apply
See all jobs
This is one of many
415,275 more open roles from verified company boards, updated every day.