Actively hiring
Natural Language Processing
LLM & Generative AI
Artificial Intelligence
AI Evaluation & Observability
As featured in
View 18 jobs
Follow
Overview
News
Technologies
Salaries
Products
People
Growth
Offices
Jobs
Financials
Funding and backers
Team
11-50
employees, estimated range
Hiring now
18
open roles
Momentum
25/100
a new city, in the news
Last sign of life
this month
a posting or a news item
The story so fareach stage dated by its source: the company's releases, the press, registers, filings
Aug 17
On the Hacker News front page Qwen3.8 27B scores 52 on Artificial Analysis · 381 points
Jun 18
Launched on Hacker News Show HN: AA-Briefcase: a frontier knowledge work evaluation
Jun 17
On the Hacker News front page GLM-5.2 is the new leading open weights model on Artificial Analysis · 916 points
Tractionpublic signals, read monthly
Web presence
top 0.29%
of 133.2M domains by links, Common Crawl
Investors in similar startupsfunds that back the companies most like this one

backs LLM Stats, Openbenchmarks

backs HoneyHive
Similar startupsby what they do, their industry and backers

The AI Leaderboard — independent rankings of GPT, Claude, Gemini, Llama, DeepSeek and 300+ AI models by intelligence, speed and price…

Independent research lab focused on reasoning efficiency, benchmarks, and compilers for large language models.

HoneyHive is an AI agent observability and evaluation platform. Teams trace what their agents did, evaluate every output, and improve…

We provide highly specialised integration services, and AI model evals for global OEMs. Supported by 100,000+ domain experts.

Testing platform to accelerate prompt iteration and reduce developer bottlenecks. Built for agentic workflows, structured outputs, and…

Monitoring and evaluation tools for the next generation of AI systems. from research to production.
News
MiMo-v2.6-Flash: on intelligence/price Pareto frontier
Analysis of Xiaomi's MiMo-V2.6-Flash and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.
Read more
Report
Claude Opus 5.5とGPT-6 Solの大幅値下げ 安くなったもの、変わらなかったもの、見えないもの山内怜史 Sun AIストラテジスト & ビジネスデザイナー
序章:AnthropicとOpenAIは、同じ9月22日に値下げを発表した 2026年9月22日16時31分UTC、日本時間では23日1時31分、AnthropicはClaude Opus 5.5を発表した Claude Opus 5.5 is available today. https://t.co/lSCmTYw9eJ - Anthropic (@AnthropicAI) September 22, 2026 Anthropicが発表した100分後の18時12分日
Read more
Report
Mercury 2.5 LLM hits 770 tokens per second
Analysis of Inception's Mercury 2.5 and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.
Read more
Report
GPT-6 Sol (Max) Intelligence, Performance and Price Analysis
Analysis of OpenAI's GPT-6 Sol (max) and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.
Read more
Report
GPT-6 Luna (Max) Intelligence, Performance and Price Analysis
Analysis of OpenAI's GPT-6 Luna (max) and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.
Read more
Report
Claude Opus 5.5 (high with fallback) - Intelligence, Performance & Price Analysis Artificial Analysis
Analysis of Anthropic's Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.
Read more
Report
Claude Opus 5.5 Intelligence, Performance and Price Analysis (Max)
Analysis of Anthropic's Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback) and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.
Read more
Report
MiMo-v2.6-Pro: Intelligence, Performance and Price Analysis
Analysis of Xiaomi's MiMo-V2.6-Pro and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.
Read more
Report
GPT-6 Sol and Luna push the cost efficiency frontier
GPT-6 Sol and Luna push the cost efficiency frontier by halving cost relative to GPT-5.6 Sol and Luna. Intelligence Index and Coding Agent Index scores remain level with GPT-5.6, with progress in some evaluations and regressions in others
Read more
Report
Grok 4.7 Intelligence, Performance and Price Analysis
Analysis of SpaceXAI's Grok 4.7 (xhigh) and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.
Read more
Report
Show more news
Tech stack
Technologies named in job postings at Artificial Analysis over the last 12 months, by layer, with the share of postings for each.
Technologies in use
57
Postings analysed
19
Stack modernity
82/100AI adopter
Languages
Python95%
JavaScript11%
Node JS11%
TypeScript11%
Frameworks & libraries
PyTorch26%
CUDA11%
Next.js11%
React.js11%
Databases
PostgreSQL11%
Cloud & platforms
Anthropic100%
GitHub100%
Hugging Face100%
OpenAI100%
AI & ML
LLM11%
Midjourney11%
Kling5%
Sora5%
Sign in to see how your skills match this stack
Sign In
Full Artificial Analysis tech stack
57 technologies · by team · pay by technology · changes · similar stacks
Growth
Hiring Momentum
52/100
Stable
Open positions
18
0 opened / 1 closed in 30 days
Median time-to-fill
59 days
faster than 26% of the market
ATS activity
Every ~3 hours
Today
Hiring Dynamics
+100%
Hiring Focus
The percentage next to each role is its share of the company's job openings over the last 90 days; the arrow shows the shift versus the previous period.
AI/ML
59% ▼
Product
18% ▼
Backend
12% ▼
HR
6% ▼
Solutions
6% ▼
Activity Timeline
New location: Australia
Sep 2026
Added Machine Learning to stack
Sep 2026
Offices
Where the company hires and what each office is for
Cities
2
Hiring now
2
Countries
2
Busiest office
San Francisco
Loading the map
Headquarters
Builds here
Office with open vacancies
Tap a pin for the office
San Francisco
United States
Engineering / R&D
16 open roles 15 posted in 90 days 69% engineering 10-50 est. staff
hires for AI/ML, Product, Solutions
Melbourne
Australia
Engineering / R&D
3 open roles 3 posted in 90 days 67% engineering 1-10 est. staff
hires for AI/ML, Backend, Solutions
Jobs
≈ $159k – $293k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • San Francisco
Python
JavaScript
TypeScript
Node JS
PostgreSQL
Supabase
AI/ML
PyTorch
OpenAI
Anthropic
Hugging Face
Edge AI
Frontend
Next.js
React.js
DevOps
Vercel
CI/CD
GitHub
Apply
≈ $170k – $323k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • San Francisco
Python
PyTorch
OpenAI
Anthropic
Hugging Face
Edge AI
DevOps
GitHub
Apply
≈ $192k – $392k per year (Estimated) • In office • Full-Time • 3+ years exp • San Francisco
Python
OpenAI
Anthropic
Hugging Face
Edge AI
Machine Learning
DevOps
GitHub
Apply
≈ $117k – $257k per year (Estimated) • In office • Full-Time • 3+ years exp • San Francisco • Melbourne
Python
LLM
Tokenization
OpenAI
Anthropic
Hugging Face
DevOps
Vercel
Datadog
Cloudflare
GitHub
Management
Slack
Stripe
Apply
≈ $193k – $395k per year (Estimated) • In office • Full-Time • 3+ years exp • San Francisco
Python
Groq
vLLM
CUDA Toolkit
TensorRT
Cerebras
CUDA
OpenAI
Anthropic
Hugging Face
TPU
Edge AI
DevOps
GitHub
Apply
≈ $194k – $396k per year (Estimated) • In office • Full-Time • 3+ years exp • San Francisco
Python
vLLM
AI Agents
SGLang
TensorRT
TensorRT-LLM
Together AI
Cerebras
OpenAI
Anthropic
Hugging Face
CoreWeave
Edge AI
Baseten
DevOps
AWS
AWS Lambda
GitHub
Apply
≈ $190k – $388k per year (Estimated) • In office • Full-Time • 3+ years exp • San Francisco
Python
AI Agents
NLP
LLM
OpenAI
Anthropic
Hugging Face
LLM Evaluation
Edge AI
DevOps
GitHub
Apply
≈ $192k – $393k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • San Francisco
Python
OpenAI
Anthropic
Hugging Face
Edge AI
DevOps
GitHub
Apply
≈ $190k – $388k per year (Estimated) • In office • Full-Time • 3+ years exp • San Francisco
Python
ElevenLabs
AI Agents
OpenAI
Anthropic
Hugging Face
Text-to-Speech
Deepgram
Edge AI
Voice Agents
DevOps
GitHub
Apply
≈ $122k – $237k per year (Estimated) • In office • Full-Time • 3+ years exp • San Francisco
Python
Midjourney
TensorFlow
NumPy
PyTorch
Runway
OpenAI
Anthropic
Hugging Face
Edge AI
DevOps
GitHub
Analytics
Matplotlib
Apply
≈ $192k – $393k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • San Francisco
Python
Databricks
AI/ML
OpenAI
Anthropic
Hugging Face
Scale AI
Edge AI
DevOps
Vercel
Datadog
GitHub
Management
Stripe
Apply
≈ $127k – $246k per year (Estimated) • In office • Full-Time • 3+ years exp • San Francisco
Python
ElevenLabs
AI Agents
OpenAI
Anthropic
Hugging Face
Text-to-Speech
Deepgram
Edge AI
Voice Agents
DevOps
GitHub
Apply




