1,434,312open jobs
83,785companies
217,826added this week
Browse all
Salary
$205k – $250k per year
Location
Hybrid (Boston, Washington, San Francisco, United States)
Seniority
Staff · 8+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 9, 2026. First seen by Alion on Oct 6, 2026. Code Metal scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Code Metal is a Boston-based AI software company that builds verifiable code translation tooling, porting, optimising and proving the correctness of software that must run on specific edge and embedded hardware in defense, aerospace, automotive, semiconductor and robotics programs. Founded in 2023 with offices in Boston and San Francisco, it lists the US Air Force, L3Harris, RTX and Toshiba as customers and is backed by Accel, Salesforce Ventures, B Capital, Shield Capital, RTX and other investors. Its hiring spans code transpilation and RTL or high-level synthesis engineering, modeling and simulation programmers, defense deployment strategists and go-to-market leadership.

About Code Metal

Code Metal is the leader in automated software engineering you can trust. As AI writes more of the world's code, the bottleneck in software has shifted from writing code to verifying it works, and AI cannot verify its own work with certainty. Code Metal takes a fundamentally different approach: constrain AI to what it does reliably, verify every step independently of the model using formal methods, and keep engineers in the loop on the decisions that matter. The result isn't code that probably works - it's code that is provably correct, with auditable proof. Customers including the U.S. Air Force, L3Harris, RTX, and Toshiba use Code Metal to modernize legacy code, optimize performance on real hardware, and move prototypes to production, fast. Founded in 2023 with offices in Boston and San Francisco, Code Metal is funded by Accel, Salesforce Ventures, B Capital, Smith Point Capital, J2 Ventures, Shield Capital, Overmatch, RTX, and others.

Learn more at codemetal.ai.

The Role

Code Metal's engineering teams are building AI-driven code transpilation and AI-enabled mission planning and wargaming. Both need the same foundations: models to serve, agents to run, context to manage, and results to measure. Our AI Platform team builds those foundations.

As a Staff AI Platform Engineer, you'll be the technical lead of this new four-person team. You'll architect and build the AI enablement stack our engineers depend on, from GPU inference serving and a model gateway up through agent harnesses, context engineering, observability, and AI experimentation management. It starts as an internal platform, but we're building it to product standard.

This is an engineering role first. Most of your time goes to designing, building, and operating production systems. You'll also need solid data science and AI research fundamentals: you'll work closely with our Applied AI Research team, and you'll sometimes run experiments yourself when a platform decision needs evidence.

Core Responsibilities

  • Set the technical direction and architecture for Code Metal's AI platform and lead the team building it. Own the design docs and RFCs, help with build-vs-buy decisions, and mentor the team.

  • Deploy, benchmark, and tune production inference for open-weight models on vLLM, SGLang, and TensorRT-LLM.

  • Own the model gateway that teams use to reach self-hosted and commercial models, with consistent auth, routing, failover, quotas, and cost attribution.

  • Design reusable agent harnesses and orchestration primitives that product teams can compose into reliable, verifiable workflows instead of rebuilding them for each product.

  • Build context-engineering services for memory, retrieval, and data discovery, so agents get the right information within their context and cost budgets.

  • Instrument the stack end to end with OpenTelemetry traces and service metrics, and build the experiment-tracking and artifact layer that lets engineers and researchers reproduce and compare results.

  • Design for productization from day one (multi-tenancy, versioned APIs, security, and deployment in customer and air-gapped environments), and partner with Applied AI Research, product teams, and DevOps so the platform stays aligned with what they need.

Required Qualifications

  • Production-grade Python and strong platform engineering fundamentals: API and service design, distributed systems, containers and Kubernetes, CI/CD, and testing.

  • Shipped production agentic systems, with a clear sense of where they break and how to make them reliable.

  • Experience with context engineering: retrieval-augmented generation, embeddings, vector or hybrid search, and memory for agents, ideally over code or large technical corpora.

  • Experience instrumenting services (for example, with OpenTelemetry tracing and metrics) and operating AI services against SLOs.

  • Solid data science and AI research fundamentals: how transformers and LLM inference work, experiment design, benchmarking, and model evaluation. Working familiarity with PyTorch and Hugging Face, and experience fine-tuning, evaluating, or serving language models.

  • Staff-level technical leadership: owned architecture across multiple systems or teams, written design docs and RFCs, turned ambiguous needs from several internal customers into a roadmap, and mentored engineers.

Preferred Qualifications

  • Highly desirable: Production experience running an LLM gateway or proxy such as SMG or Bifrost, or equivalent experience building an API gateway, including routing, auth, rate limiting, quotas, failover, and cost attribution.

  • Hands-on experience deploying and tuning LLM inference engines such as vLLM, SGLang, or TensorRT-LLM on GPU infrastructure, with measurable gains in throughput, latency, or cost per token.

  • Familiarity with inference optimization: speculative decoding, prefix caching, tensor/pipeline/expert parallelism, disaggregated prefill and decode, GPU profiling.

  • Experience building evaluation harnesses for LLMs and agents, and experiment-tracking or artifact systems such as MLflow or Weights & Biases.

  • Experience taking an internal platform to an external product: multi-tenancy, SDKs, versioned APIs, and documentation.

  • Experience deploying AI systems on-prem or in air-gapped or classified environments, or in regulated domains such as defense or aerospace.

Experience Level

Typically 8+ years of software engineering experience, including 4+ years building and operating ML or LLM systems in production, with demonstrated Staff-level scope and impact. Equivalent depth of experience is valued over a rigid year count.

Benefits

  • Pay depends on experience, but we strive to be at the upper end of the salary range

  • Health care plan with 100% premium coverage, including medical, dental, and vision

  • 401k with 5% matching

  • Paid Time Off (uncapped vacation, plus sick and public holidays)

  • Flexible hybrid or remote work arrangement

  • Relocation assistance for qualifying employees

Wage Transparency - The salary range for this role is not a guarantee of compensation or salary, as the final offer amount may vary based on factors including, but not limited to, individual proficiency, skills, experience, and location.

We are an equal opportunity employer. US Citizenship may be required for certain project assignments involving security clearance.

Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,434,312 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Boston
AI Engineer IV 7 months ago
$150k – $271k per year • Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Mountlake Terrace
AI/ML
Diffusion Models
MLX ML
TensorFlow
PyTorch
LLM
RAG
DevOps
Azure
Management
Agile
Apply
Principal AI Engineer 3 hours ago
$230k – $280k per year • Equity • Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Palo Alto
Python
AI/ML
AI Agents
Mistral
LLM
OpenAI
Anthropic
LLM Guardrails
Agentic Workflows
Apply
≈ $152k – $316k per year (Estimated) • In office • TS/SCI • Full-Time • Master's Degree • Herndon
AI/ML
Computer Vision
Machine Learning
Apply
$201k – $258k per year • Hybrid • 5+ years exp • PhD • San Francisco
Python
AI/ML
DeepSpeed
Quantization
JAX
Multimodal AI
Knowledge Distillation
Diffusion Models
Accelerate
Transformers
PyTorch
Ray
In-Context Learning
Mixture of Experts
Hugging Face
Megatron-LM
FSDP
GraphRAG
Knowledge Graph
Model Distillation
Machine Learning
Apply
$179k – $232k per year • Equity • In office • Full-Time • Bachelor's Degree • Mountain View • Bellevue
Python
Databases
Apache Kafka
AI/ML
Ray Serve
Spark
Prefect
PyTorch
Ray
KServe
Triton
Feature Store
Machine Learning
DevOps
GCP
Prometheus
AWS
Kubernetes
Grafana
Game Dev
Unity
Apply
$44k – $88k per year • In office • Internship • Bachelor's Degree • Fort Wayne
Python
C++
Apply
Sr Systems Analyst 6 hours ago
≈ $76k – $147k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Anaheim
Python
Java
SQL
C#
C#
.NET
Databases
Snowflake
Oracle
Analytics
SAP BusinessObjects
Management
Agile
ITIL
Apply
$144k – $193k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Orlando • Anaheim • Seattle • Glendale
Python
Java
Java
Spring Boot
Databases
Snowflake
DynamoDB
Apache Kafka
Amazon Redshift
AI/ML
Spark
Scikit-learn
Computer Vision
NLP
TensorFlow
PyTorch
Amazon SageMaker
Recommender Systems
Machine Learning
DevOps
CI/CD
AWS
Docker
Kubernetes
Amazon EC2
Amazon S3
Apply
AI Engineer 6 hours ago
$100k – $130k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • United States
Python
Java
TypeScript
AI/ML
Copilot
Cursor
Claude
Claude Code
Model Context Protocol
AI Agents
AWS Bedrock
LLM
OpenAI Codex
OpenAI Agents SDK
LLM Guardrails
Machine Learning
DevOps
Terraform
CI/CD
Git
AWS
AWS Lambda
Amazon S3
IAM
API Gateway
Apply
≈ $35k – $73k per year (Estimated) • Remote (Canada) • Full-Time • 1+ year exp • Toronto
Python
SQL
SAS
Analytics
Power BI
Apply
$170k – $210k per year • Hybrid • Full-Time • 4+ years exp • Boston • San Francisco • Washington
Python
AI/ML
Weights & Biases
Model Context Protocol
vLLM
MLFlow
Fine-tuning
Embeddings
AI Agents
SGLang
TensorRT
TensorRT-LLM
PyTorch
LLM
RAG
Hybrid Search
Hugging Face
Context Engineering
LLM Evaluation
Speculative Decoding
DevOps
OpenTelemetry
CI/CD
Kubernetes
Platform Engineering
API Gateway
Apply
≈ $137k – $302k per year (Estimated) • Hybrid • Full-Time • Bachelor's Degree • Boston • San Francisco
Python
AI/ML
vLLM
Fine-tuning
Embeddings
Reinforcement Learning
Knowledge Distillation
Function Calling
AI Agents
Transformers
PyTorch
LLM
RAG
Synthetic Data
Hugging Face
Post-training
Human-in-the-Loop
Structured Outputs
LLM Guardrails
Agentic Workflows
Multi-Agent Systems
Tool Use
Model Distillation
Machine Learning
Apply
≈ $132k – $292k per year (Estimated) • Hybrid • Top Secret • Full-Time • Master's Degree • Boston • San Francisco
Python
Rust
C++
Python
Hypothesis
DevOps
CI/CD
Apply
≈ $59k – $120k per year (Estimated) • In office • Top Secret • Full-Time • 5+ years exp • Boston • Washington • San Francisco
Management
Slack
Apply
$153k – $183k per year • Hybrid • Full-Time • 8+ years exp • Boston • Washington • San Francisco
Python
Rust
C++
AI/ML
AI Agents
Agentic Workflows
DevOps
CI/CD
Linux
Apply
≈ $128k – $282k per year (Estimated) • Equity • Hybrid • Secret • Full-Time • Bachelor's Degree • Boston
Python
AI/ML
OpenCV
YOLO
Scikit-learn
Computer Vision
TensorFlow
PyTorch
Anomaly Detection
Machine Learning
DevOps
Git
Apply
$53k – $60k per year • Hybrid • Internship • Bachelor's Degree • Boston
Design
AutoCAD
Apply
$53k – $60k per year • Hybrid • Internship • Bachelor's Degree • Boston
Design
AutoCAD
Apply
Sales Executive 10 hours ago
Hybrid • Bachelor's Degree • Boston
Marketing
Salesforce
Apply
$116k – $184k per year • Equity • Remote (United States) • Full-Time • 15+ years exp • Bachelor's Degree • Boston
DevOps
Self-Healing
QA
Rest-Assured
Apply
See all jobs
This is one of many
1,434,312 more open roles from verified company boards, updated every day.