Overview
News
Salaries
Products
People
Growth
Offices
Financials

Funding and backers

Backers
1
Y Combinator
Momentum
0/100
no fresh signals
Last sign of life
this month
a posting or a news item
Tractionpublic signals, read monthly
Web presence
top 20%
of 133.2M domains by links, Common Crawl
Backed byfrom the funds' portfolios and accelerator catalogs
Similar startupsby what they do, their industry and backers
DeepInfraSeries B $107M · 4 mo ago • Founded 2022 • 501-1000 employees • 16 roles
DeepInfra is a company founded in 2022 that runs a low-cost inference cloud for open machine learning models. Developers call hosted…
Novita AIFounded 2023 • 1-10 employees • 2 roles
Novita AI provides serverless graphics processing units and hosted model APIs. Its platform runs open source image and language models on…
hostedVenture round $18.7M · 6 mo ago • Founded 2024 • 11-50 employees
Build profitable AI clouds. Monetize GPU infrastructure. Ultra efficient GPUaaS, optimized GPU utilization, capacity sharing, 5x more…
Computable GPU Index (CGI) is a USD price per GPU-hour, computed from the published on-demand rental rates of a fixed panel of providers.
Hydra HostVenture round $70M · 6 mo ago • Founded 2021
Hydra Host offers Bare Metal GPU solutions for AI, HPC, and big data. Enjoy autonomy, privacy, and scalable performance options. Get in…
Andromeda11 roles
Andromeda connects AI teams with high-performance compute fast, at scale, and on terms that work. Buy, sell, and operate gpu clusters…

Overview

OpenRelay provides hosted model inference and dedicated GPU VMs, with per-token and usage-based pricing.

News

Blog 1 day ago
orl 0.4: GPU pods, VM sizing, and spend in the OpenRelay CLI
orl 0.4 adds GPU pods, VM sizing checks before launch, per-VM burn rates, and expedited batch inference. Install it and rent a GPU from your terminal.
Read more
Report
Blog 1 month ago
AI Best Practices: Cutting LLM Pipeline Costs Without Losing Quality
Two changes that cut most LLM pipeline bills 3-10x: right-size the model so a 20B screens the easy majority, and cache the shared prefix so repeated system prompts bill at one tenth the rate.
Read more
Report
Blog 3 months ago
OpenRelay is Backed by Y Combinator
We're building the CDN of inference: a distributed GPU network that makes fast, affordable, fault-tolerant AI compute available to everyone. Here's why, and what comes next.
Read more
Report
Blog 5 months ago
How One Team Cut $4,000/mo in Vercel Build Costs with OpenRelay Runners
Replacing Vercel's build infrastructure with OpenRelay self-hosted GitHub runners saved $4,000/month in build minutes: with faster builds and zero config overhead.
Read more
Report
Blog 8 months ago
The Environmental Case for Distributed GPU Computing
Why reusing existing consumer GPUs for AI inference is greener than building new data centers. The environmental argument for distributed networks.
Read more
Report
Pro access
Upgrade to see all 30 mentions
Upgrade to a paid plan to read every media mention of this company - funding news, awards, product launches and press releases from all the outlets writing about it.
Every media mention and press release
Funding news, awards and product launches
Fresh coverage from every outlet writing about the company
Upgrade now
Cancel anytime. Secure checkout. Instant activation.