771,423open jobs
48,517companies
117,598added this week
Browse all
Salary
≈ $166k – $302k per year (Estimated)
Location
In office (San Mateo)
Seniority
Senior · 6+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 25, 2026. First seen by Alion on Sep 24, 2026.

Overview
Company
Impact
Profile match
Deploy Physical AI agents powered by Newton foundation model. Fuse sensor data, automate decisions, and build intelligent solutions for manufacturing, construction, and more.

About Job

At Archetype AI, we’re building the world’s first physical AI platform to bring artificial intelligence into the real world. Our foundation model, Newton, understands the physical world through objective sensor data and generates real-time insights into complex physical behaviors, from industrial machinery and systems to wearable devices and smart environments.

Formed by a high-caliber team from Google and backed by one of Silicon Valley’s most renowned venture funds, Archetype AI is in a Series A phase and rapidly advancing its technology for the next big leap. This is a unique opportunity to join an exciting, fast-growing AI team based in the heart of Silicon Valley.

About the Role

You will own the serving path for Newton and related multimodal models. Much of our inference stack is Rust-native: model nodes in our agent runtime, built on Rust ML stacks (candle, Burn) with custom GPU kernels, plus the routing layer that streams real-time inference to GPU nodes. You will drive GPU utilization, numerical precision, and low-latency serving from the kernel up.

What You'll Own

  • Build and own model nodes in our Rust inference runtime: loading, warmup, batching, streaming, GPU memory pools.

  • Optimize kernels and the GPU path: custom CUDA kernels, mixed precision, quantization, parity against research.

  • Own inference routing and serving: streaming path API to GPU node, request batching, SLO-backed latency and cost.

  • Productionize research checkpoints: export, compilation, quantization, parity evals, and rollout.

  • Build the observability inference needs: latency histograms, GPU metrics, OOM signatures, replayable traces.

Key Qualifications

  • 6+ years software engineering, several of them in ML systems, inference, or high-performance GPU computing.

  • Has owned a production serving path end to end, not only benchmarked models.

  • Expert in PyTorch with models shipped to production; knows what will be slow before the profiler runs.

  • Strong CUDA or equivalent GPU depth: memory hierarchy, occupancy, Nsight or equivalent profiling.

  • Rust or C++ alongside Python, Linux performance, production ops; ready to work in Rust daily.

  • Works with researchers: can translate an architecture change into serving work.

Nice to Have

  • Rust ML stacks: candle, Burn, or comparable GPU compute in Rust.

  • Custom kernels and compiler stacks: Triton, CUTLASS, TorchInductor, TensorRT.

  • Quantization and mixed precision in production with a numerical-correctness suite.

  • Multimodal, video, embedding or time-series serving, not only decoder-only chat LLMs.

  • High-performance serving stacks (vLLM, SGLang, TensorRT-LLM): continuous batching, paged KV cache.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
771,423 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
San Mateo
AI Platform Engineer 1 month ago
$184k – $215k per year • Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Rockville
Databases
PostgreSQL
Redis
DynamoDB
OpenSearch
Amazon Aurora
Amazon DocumentDB
AI/ML
Claude Code
LLM
DevOps
Terraform
Datadog
GitLab CI
CI/CD
AWS
Docker
Amazon EC2
Amazon S3
Amazon ECS
Amazon CloudWatch
Linux
Apply
$180k – $205k per year • Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Rockville
Python
JavaScript
TypeScript
Node JS
Python
FastAPI
Databases
PostgreSQL
Redis
Neo4j
Apache Iceberg
Delta Lake
DynamoDB
OpenSearch
Amazon Aurora
Apache Hudi
Amazon Neptune
Amazon DocumentDB
AI/ML
LangGraph
LangChain
Claude Code
LlamaIndex
Model Context Protocol
Function Calling
AI Agents
Arize Phoenix
DeepEval
Ragas
AWS Bedrock
LLM
RAG
Hallucination
AWS Strands Agents
LLM Evaluation
Agentic Workflows
Tool Use
Frontend
Tailwind CSS
React.js
Vite
DevOps
Rest API
Terraform
Datadog
WebSockets
GitLab CI
CI/CD
Git
AWS
Docker
AWS Lambda
Amazon EC2
SLI/SLO/SLA
GitLab
Amazon S3
IAM
Amazon ECS
Amazon CloudWatch
Amazon Kinesis
Amazon EventBridge
API Gateway
Linux
Apache HTTP Server
Cybersecurity
FedRAMP
Analytics
ETL/ELT
AWS Glue
Management
Confluence
Jira
Apply
$160k – $205k per year • In office • Full-Time • 5+ years exp • Mountain View
Python
Java
C++
Scala
Databases
Aerospike
AI/ML
AI Agents
Machine Learning
DevOps
Docker
Kubernetes
GitHub
Apply
$250k – $600k per year • In office • Full-Time • New York • Geneva • London • Paris
AI/ML
World Models
Apply
≈ $180k – $326k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • San Francisco • New York
AI/ML
Groq
vLLM
SGLang
TensorRT
TensorRT-LLM
TGI
LLM
Cerebras
TPU
AWS Trainium
Speculative Decoding
KV Cache
Apply
≈ $25k – $62k per year (Estimated) • In office • Noida
Python
Go
Java
DevOps
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Cybersecurity
GDPR
HIPAA
Apply
Senior Data Engineer 5 hours ago
≈ $122k – $210k per year (Estimated) • Remote (United States) • Full-Time
Python
SQL
Databases
PostgreSQL
Snowflake
Databricks
Google BigQuery
Amazon Redshift
BigQuery
AI/ML
Claude Code
dbt
Great Expectations
LLM
DevOps
Terraform
AWS
Cybersecurity
GDPR
HIPAA
Analytics
Power BI
Fivetran
Data Vault
Apply
≈ $23k – $44k per year (Estimated) • In office • Saint Petersburg
C++
DevOps
Windows
Cybersecurity
Dallas Lock
Apply
≈ $33k – $70k per year (Estimated) • In office • Moscow
Python
AI/ML
GigaChat
Transformers
PyTorch
LLM
Mixture of Experts
Apply
$118k – $160k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • Chelmsford
C++
MATLAB
MATLAB
Simulink
Robotics
EtherCAT
Sensor Fusion
Apply
AI Researcher 1 day ago
≈ $163k – $295k per year (Estimated) • Hybrid • Full-Time • 8+ years exp • San Mateo
Python
AI/ML
Multimodal AI
PyTorch
Self-Supervised Learning
Contrastive Learning
Physical AI
Apply
≈ $158k – $342k per year (Estimated) • Hybrid • Full-Time • San Mateo
AI/ML
Multimodal AI
LLM
Time Series Forecasting
Physical AI
Apply
≈ $147k – $263k per year (Estimated) • Hybrid • Full-Time • 8+ years exp • San Mateo
Python
Go
Rust
C++
AI/ML
Physical AI
DevOps
Rest API
GCP
Azure
CI/CD
AWS
Kubernetes
IoT
MQTT
OPC UA
Apply
≈ $133k – $278k per year (Estimated) • In office • Full-Time • San Mateo
AI/ML
Time Series Forecasting
Physical AI
Apply
Remote (Japan) • Full-Time
AI/ML
Time Series Forecasting
Physical AI
Apply
Senior ML Engineer 5 hours ago
$100k – $160k per year • Equity 0.3–0.3% • Remote (location not specified) • Full-Time • 11+ years exp • Bachelor's Degree • San Mateo
AI/ML
OpenCV
Fine-tuning
Computer Vision
Machine Learning
Apply
$100k – $160k per year • Equity 0–0.2% • In office • Full-Time • 11+ years exp • Bachelor's Degree • San Mateo
Python
C++
C++
TensorFlow C++
AI/ML
OpenCV
Fine-tuning
TensorFlow
Machine Learning
Apply
Sales Engineer 5 hours ago
$90k – $110k per year • In office • Full-Time • 6+ years exp • Bachelor's Degree • San Mateo
Python
Management
Confluence
Apply
$48k – $56k per year • In office • Internship • Bachelor's Degree • San Mateo
PHP
PHP
WordPress
Design
Adobe Photoshop
Adobe After Effects
Marketing
HubSpot
YouTube
Apply
$90k – $110k per year • Remote (United States) • Full-Time • 3+ years exp • San Mateo
Apply
See all jobs
This is one of many
771,423 more open roles from verified company boards, updated every day.