rate card/llm-inference
X402 RATE CARD · LLM INFERENCE

What does LLM inference cost an agent per call over x402?

The median LLM inference provider on x402 charges $0.02 per call. The middle half of the market sits between $0.005 and $0.1, based on 933 distinct priced offers from 152 providers. Updated 2026-08-05.

Chat completions, embeddings and model routing billed per request in USDC instead of a prepaid credit balance.

MEDIAN / CALL
$0.02
provider-weighted
TYPICAL RANGE
$0.005–$0.1
p25 – p75
PROVIDERS
152
933 priced offers
PRICE VERIFIED LIVE
43%
401 of 933 offers

Where a price would land

Each rung is the provider-weighted percentile across this category. Pick the position you want to hold, then read the price.

p10 — undercut the field$0.001
p25 — value position$0.005
median — market rate$0.02
p75 — premium position$0.1
p90 — top of market$0.5

Price distribution

Distinct priced offers per bracket — templated route families collapsed so one product doesn’t count hundreds of times.

≤ $0.001
123
$0.001–$0.01
490
$0.01–$0.05
176
$0.05–$0.10
55
$0.10–$0.50
63
$0.50–$1
11
$1–$10
10
> $10
5

⚠ concentration: llm402.ai publishes 40% of the distinct offers in this category. The provider-weighted median above corrects for this; the raw per-route median is $0.008.

Provider comps

What each seller actually publishes in this category, ranked by catalogue size.

PROVIDERMEDIANRANGEOFFERSNETWORKS
llm402.ai
llm402.ai
$0.003076$0.000614$1.35168373 / 387 routesBase
BlockRun.AI
blockrun.ai
$0.0095$0.002$0.2645175Base
x402node
api.x402node.dev
$0.005$0.003$0.0229Base · Solana
DJD Agent Score
djdagentscore.dev
$0.15$0.01$9921Base
GoCreative AI
api.gocreativeai.com
$0.01$0.003$0.2517 / 18 routesBase
the-stall
the-stall.intuitek.ai
$0.059$0.034$0.3716Base
myceliasignal
api.myceliasignal.com
$0.02$0.02$0.0515Base
agentutility
x402.agentutility.ai
$0.01$0.005$0.0813Base
gpt55
gpt55.558686.xyz
$0.006$0.002$0.02912 / 13 routesBase · Solana
strale
api.strale.io
$0.0324$0.0216$0.1628Base
webbersites
api.webbersites.com
$0.001$0.001$0.026Base
AgentBodega
agentbodega.store
$0.01$0.005$0.056Base
vape-x402
vape-x402.vapex402.workers.dev
$0.02$0.01$0.16Base
x402render
x402render.vercel.app
$0.04$0.006$0.066Base
glianalabs
api.glianalabs.com
$0.048$0.037632$0.2949126Base · Solana
orbisapi
orbisapi.com
$0.001$0.001$0.0035Base
x402
x402.ottoai.services
$0.002$0.001$0.15Polygon · Base · Solana
news-summarizer-x402.vercel.app
news-summarizer-x402.vercel.app
$0.005$0.005$0.0055Base
OpenRouter
api.paysponge.com
$0.01$0.01$0.015Base · Solana
promptqualityscore
promptqualityscore.com
$0.025$0.001$0.15Base

Live routes across the price range

Real endpoints sampled from the cheapest to the most expensive offer in this category.

true402 LLM Router
GET https://true402.dev/api/v1/chat/completions
true402.dev$0.0001
Meta-Llama-3-8B-Instruct-Lite
POST https://llm402.ai/v1/chat/completions/Meta-Llama-3-8B-Instruct-Lite
llm402.ai$0.001
gpt-oss-120b
POST https://llm402.ai/v1/chat/completions/gpt-oss-120b
llm402.ai$0.00169
cogito-v2-1-671b
POST https://llm402.ai/v1/chat/completions/cogito-v2-1-671b
llm402.ai · ×2$0.002816
API Gateway: Groq LLM API
GET https://httpay.xyz/api/gateway/groq/%7B*path%7D/*
Alfred's httpay.xyz$0.003
x402node$0.005
Purchase: Content Summarization
GET https://api.the402.ai/v1/services/svc_f6a85cc2496845c8/purchase
the402$0.008
https://blockrun.ai/api/v1/surf/search/fund
GET https://blockrun.ai/api/v1/surf/search/fund
blockrun$0.0095
claude-haiku-4-5
POST https://llm402.ai/v1/chat/completions/claude-haiku-4-5
llm402.ai · ×2$0.011264
command-a
POST https://llm402.ai/v1/chat/completions/command-a
llm402.ai$0.022528
https://api.myceliasignal.com/oracle/defi/metrics
GET https://api.myceliasignal.com/oracle/defi/metrics
myceliasignal$0.05
gpt-4
POST https://llm402.ai/v1/chat/completions/gpt-4
llm402.ai$0.135168

Methodology

Every figure is built from prices providers publish for their own endpoints, captured by the TOLL·402 discovery crawl and rebuilt 2026-08-05. Where a strict HTTP 402 check succeeded we use the price the endpoint actually quoted rather than its catalogue entry — 401 of the 933 offers here are priced from a live quote. That matters: across the corpus, roughly one in twenty-seven catalogued prices disagrees with the live quote, some by more than 100×.

Two corrections are then applied before any median is taken. First, templated route families — one endpoint per card, per ticker, per city — are collapsed into distinct offers keyed by provider, title and price, so a 900-route family counts as one price point. Second, each provider contributes the median of its own offers exactly once, so a single bulk seller cannot move the benchmark.

Prices are what providers publish in their 402 quote. They are not evidence of demand, and a confirmed quote is not a settled payment. A category is only published once at least 4 distinct providers are observed in it.

FAQ

How much should I charge for a LLM inference API on x402?

Across 152 providers publishing LLM inference routes, the median seller charges $0.02 per call. The middle half of the market sits between $0.005 and $0.1. Pricing at or just below the median is the least friction for a new listing; above $0.5 you are at the top of the observed market and need a clear capability reason.

Where do these numbers come from?

TOLL·402 crawls the public x402 discovery surface and records the price each endpoint publishes in its HTTP 402 quote. This page covers 933 distinct priced offers from 152 providers, last rebuilt 2026-08-05. Prices are what sellers publish — not settled payments and not evidence of demand.

Why is the median different from the average route price?

Because the corpus is concentrated. A single provider can register thousands of near-identical routes and drag any per-route average toward its own list price. We collapse templated route families into distinct offers, then let each provider contribute its own median exactly once — so this benchmark reflects what a typical seller charges, not what the largest seller charges.