Salary
≈ $138k – $301k per year (Estimated)
Location
Remote (United States)
Overview
Company
Impact
Profile match
AI agents for the work your team has been afraid to automate. Private by default, verifiable by design.
Locations: San Francisco or Remote
About The Role
The NEAR AI team is building decentralized and confidential machine learning infrastructure to enable user-owned AI. Our mission is to build highly scalable and efficient infrastructure for open-source AI at a global scale.
We are specifically seeking an expert in high-performance LLM serving systems and inference optimization. In this role, you will push the boundaries of how large language models are served.
What You'll Be Doing
- Architect and maintain production high-traffic LLM serving systems.
- Optimize throughput, latency, and cost for leading open-source LLMs.
What We're Looking For
- Strong hands-on experience in LLM inference, with expertise debugging and optimizing major inference engines such as SGLang, vLLM, or TensorRT.
- Deep knowledge of state-of-the-art GPU architectures, and effectively exploit them using PyTorch, Triton, CuTe, CUDA, etc.
- Proven track record in designing and maintaining end-to-end high-traffic LLM serving systems.
- Strong problem-solving skills and ability to communicate technical ideas clearly.
We'd Love If You Have
- Experience with Trusted Execution Environments (TEE).
- Active contributor to open-source LLM inference engines.
Please let us know if you require any special requirements for your interview and we'll do our best to accommodate.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,910 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Free forever. No card. Under a minute.
Your match
How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.
Recommended for you based on this role
Similar stack
Same company
San Francisco
Cyber Threat Defense Sr AI/ML Engineer
4 hours ago
$145k – $193k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Chicago • Washington • Denver • Jersey City • Boston
Python
AI/ML
AI Agents
Anomaly Detection
Embeddings
Fine-tuning
Function Calling
Hugging Face
LangChain
LLM
LLM Evaluation
LLM Guardrails
PyTorch
RAG
Red Teaming
Scikit-learn
Vertex AI
DevOps
Azure
GCP
Cybersecurity
Threat Modeling
Apply
Lead Scientist - Artificial Intelligence
6 hours ago
$98k – $164k per year • In office • Full-Time • 5+ years exp • Master's Degree • United States
Python
AI/ML
LLM
Multimodal AI
PyTorch
TensorFlow
Knowledge Graph
OpenAI
Time Series Forecasting
DevOps
GitHub
Apply
Senior Full-Stack Security/GRC Platform Engineer
4 hours ago
$87k – $130k per year • In office • Full-Time • 6+ years exp • Murray
Python
TypeScript
JavaScript
Python
Alembic
FastAPI
Pydantic
SQLAlchemy
Databases
PostgreSQL
Redis
AI/ML
Embeddings
LLM
LLM Guardrails
Ollama
RAG
Frontend
React Query
React Router
React.js
Vite
DevOps
AWS
CI/CD
Docker
Docker Compose
IAM
Terraform
Cybersecurity
FedRAMP
NIST 800-53
QA
Pytest
Apply
Senior Engineer- Agentic AI
4 hours ago
$122k – $200k per year • In office • Full-Time • 10+ years exp • Charlotte • New York
AI/ML
AI Agents
Context Engineering
Copilot
Hallucination
Human-in-the-Loop
Knowledge Graph
LLM
LLM Guardrails
LLMOps
Prompt Engineering
RAG
DevOps
Azure
CI/CD
GitHub
Kubernetes
Vector
Analytics
A/B Testing
Apply
Member of Technical Staff - Data Science
5 hours ago
≈ $187k – $378k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp
Python
AI/ML
AI Agents
Edge AI
Embeddings
LLM
NumPy
Pandas
Spark
Analytics
A/B Testing
Apply
Member of Technical Staff, Agentic Systems
2 hours ago
$130k – $500k per year • Equity • In office • Full-Time • 5+ years exp • San Francisco
AI/ML
AI Agents
Claude
Claude Code
Copilot
Cursor
Function Calling
Human-in-the-Loop
LLM Guardrails
DevOps
GitHub
Apply
VIOLENCE RESPONSE ANALYST (1823)
2 hours ago
≈ $89k – $193k per year (Estimated) • In office • Full-Time • 3+ years exp • High School Diploma • San Francisco
Apply
Staff Product Manager, AI Growth
3 hours ago
$190k – $270k per year • Remote/Hybrid • Full-Time • 7+ years exp • San Francisco
SQL
AI/ML
LLM
Recommender Systems
Analytics
A/B Testing
Marketing
Salesforce
Apply
⚡️ Founding Full Stack Software Engineer ⚡️
3 hours ago
$170k – $220k per year • Equity 1–2.8% • In office • Full-Time • 3+ years exp • San Francisco
Python
SQL
Python
Django
AI/ML
AI Agents
Context Engineering
LLM
LLM Evaluation
RAG
Apply
Data Infrastructure Engineer
4 hours ago
$170k – $230k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • New York • San Francisco
Go
Java
Kotlin
Python
Scala
Databases
Apache Kafka
Databricks
Delta Lake
Snowflake
AI/ML
Dagster
dbt
Flink
Spark
Knowledge Graph
Apply
This is one of many
368,910 more open roles from verified company boards, updated every day.

