rate card/llm-inference
X402 RATE CARD · LLM INFERENCE

What does LLM inference cost an agent per call over x402?

The median LLM inference provider on x402 charges $0.005 per call. The middle half of the market sits between $0.005 and $0.02, based on 1,426 distinct priced offers from 388 providers. Updated 2026-08-25.

Chat completions, embeddings and model routing billed per request in USDC instead of a prepaid credit balance.

MEDIAN / CALL
$0.005
provider-weighted
TYPICAL RANGE
$0.005–$0.02
p25 – p75
PROVIDERS
388
1,426 priced offers
PRICE VERIFIED LIVE
39%
553 of 1,426 offers

Where a price would land

Each rung is the provider-weighted percentile across this category. Pick the position you want to hold, then read the price.

p10 — undercut the field$0.003
p25 — value position$0.005
median — market rate$0.005
p75 — premium position$0.02
p90 — top of market$0.15

Price distribution

Distinct priced offers per bracket — templated route families collapsed so one product doesn’t count hundreds of times.

≤ $0.001
144
$0.001–$0.01
757
$0.01–$0.05
291
$0.05–$0.10
79
$0.10–$0.50
99
$0.50–$1
9
$1–$10
37
> $10
10

⚠ concentration: llm402.ai publishes 31% of the distinct offers in this category. The provider-weighted median above corrects for this; the raw per-route median is $0.006759.

Provider comps

What each seller actually publishes in this category, ranked by catalogue size.

PROVIDERMEDIANRANGEOFFERSNETWORKS
llm402.ai
llm402.ai
$0.004056$0.001$1.35168444 / 462 routesBase
BlockRun.AI
blockrun.ai
$0.006$0.002$0.2635171Base
agentutility
x402.agentutility.ai
$0.02$0.002$0.553Base
m2mcent
api.m2mcent.com
$0.1$0.01$1.530Base
x402node
api.x402node.dev
$0.005$0.003$0.0229Base · Solana
agentstools
api.agentstools.dev
$0.02$0.005$0.124Polygon · Arbitrum · eip155:480 · Base
DJD Agent Score
djdagentscore.dev
$0.15$0.01$9921Base
robbiegeorgephotography
www.robbiegeorgephotography.com
$5$5$2521Base
delx
api.delx.ai
$0.001$0.001$0.0119Base · Solana
GoCreative AI
api.gocreativeai.com
$0.01$0.003$0.2517 / 18 routesBase
Zinin M2M Hub
api.timzinin.com
$0.05$0.01$0.117Base
netintel-production-440c
netintel-production-440c.up.railway.app
$0.05$0.005$0.113Base
gpt55
gpt55.558686.xyz
$0.006$0.002$0.02911 / 12 routesBase
myceliasignal
api.myceliasignal.com
$0.02$0.02$0.0211Base
glianalabs
api.glianalabs.com
$0.043008$0.008602$0.911Base · Solana
strale
api.strale.io
$0.0324$0.0216$0.1628Base
x402
x402.ottoai.services
$0.005$0.001$0.17Polygon · Base · Solana
webbersites
api.webbersites.com
$0.01$0.001$0.057Base
x402render
x402render.vercel.app
$0.04$0.006$0.066Base
ot-intel-api
ot-intel-api.onrender.com
$0.12$0.04$0.26Base

Live routes across the price range

Real endpoints sampled from the cheapest to the most expensive offer in this category.

SYNTHORA Bridge Risk
GET https://bridgerisk.hergertsynthora.com/service
SYNTHORA$0.0008
nova-lite-v1
POST https://llm402.ai/v1/chat/completions/nova-lite-v1
llm402.ai$0.001
https://blockrun.ai/api/v1/defillama/prices/coingecko:ethereum
GET https://blockrun.ai/api/v1/defillama/prices/coingecko:ethereum
blockrun$0.002
ernie-4.5-vl-424b-a47b
POST https://llm402.ai/v1/chat/completions/ernie-4.5-vl-424b-a47b
llm402.ai$0.002816
AgentKit /api/text-summarize
GET https://44fa-2402-a00-401-c723-c52d-e4d1-3470-f49.ngrok-free.app/api/text-summarize
agentkit-x402$0.005
AgentKit /api/text-summarize
GET https://e39d-2402-a00-401-c723-c52d-e4d1-3470-f49.ngrok-free.app/api/text-summarize
agentkit-x402$0.005
gemini-3.1-flash-image (x402)
POST https://llm402.ai/v1/chat/completions/gemini-3.1-flash-image
llm402.ai$0.006759
Summarize how a site distributes content across Open Graph, feeds, socials, and ...
GET https://api.delx.ai/api/v1/x402/content-distribution-report
delx$0.01
gpt-5.4:batch
POST https://llm402.ai/v1/chat/completions/gpt-5.4%3Abatch
llm402.ai$0.016896
https://api.monpham.work/v1/terra/chat/completions
GET https://api.monpham.work/v1/terra/chat/completions
monpham$0.025
EU Accessibility Privacy Lead Scanner
POST https://api.timzinin.com/api/eu-accessibility-privacy-lead-scanner
Zinin M2M Hub$0.05
Agent sends frames/prompt -> gets animated GIF
GET https://x402.aurelianflo.com/api/do/gif/generate
x402$0.15

Methodology

Every figure is built from prices providers publish for their own endpoints, captured by the TOLL·402 discovery crawl and rebuilt 2026-08-25. Where a strict HTTP 402 check succeeded we use the price the endpoint actually quoted rather than its catalogue entry — 553 of the 1,426 offers here are priced from a live quote. That matters: across the corpus, roughly one in twenty-seven catalogued prices disagrees with the live quote, some by more than 100×.

Two corrections are then applied before any median is taken. First, templated route families — one endpoint per card, per ticker, per city — are collapsed into distinct offers keyed by provider, title and price, so a 900-route family counts as one price point. Second, each provider contributes the median of its own offers exactly once, so a single bulk seller cannot move the benchmark.

Prices are what providers publish in their 402 quote. They are not evidence of demand, and a confirmed quote is not a settled payment. A category is only published once at least 4 distinct providers are observed in it.

FAQ

How much should I charge for a LLM inference API on x402?

Across 388 providers publishing LLM inference routes, the median seller charges $0.005 per call. The middle half of the market sits between $0.005 and $0.02. Pricing at or just below the median is the least friction for a new listing; above $0.15 you are at the top of the observed market and need a clear capability reason.

Where do these numbers come from?

TOLL·402 crawls the public x402 discovery surface and records the price each endpoint publishes in its HTTP 402 quote. This page covers 1,426 distinct priced offers from 388 providers, last rebuilt 2026-08-25. Prices are what sellers publish — not settled payments and not evidence of demand.

Why is the median different from the average route price?

Because the corpus is concentrated. A single provider can register thousands of near-identical routes and drag any per-route average toward its own list price. We collapse templated route families into distinct offers, then let each provider contribute its own median exactly once — so this benchmark reflects what a typical seller charges, not what the largest seller charges.

Before you buy at this price