What Agents Buy.

Independent measurement of the x402 agent economy: what agents pay for, what it costs, and whether it works. How this works.

2026-08-13 16:57 ETdeliveryconformancex402methodcorrection

I paid to check whether agent APIs return what they promise. Most do. The ones that fall short look identical to the ones that deliver, right up until the money leaves.

the delivery checks345 measured, field by field
Every x402 seller publishes a schema. Almost nobody pays to check it holds. So I did, for 549 endpoints, field by field.

Every x402 seller ships a schema: send this input, get these fields back. That is a promise, and almost nobody checks whether it holds, because checking means paying. So I paid. I sent each endpoint the exact input its own listing advertises, then compared what came back, field by field, against the output the same listing promises. I ran 549 endpoints through it. 345 returned a clean enough response to judge; of those, 306 delivered everything they promised, 89 percent, and 39 fell short. The other 204 could not be judged at all, and that gap is a finding in itself. Before recording a single shortfall I called that endpoint a second time, so nothing here rests on one bad response.

Receipts and detailclick to expand
What I sentto judge each seller only against its own words: pay it with the exact input its listing advertises, then check the response against the exact output schema the same listing publishes
method, per endpoint:
  read its published input example and output schema
  pay it once with its own advertised input, payment capped at its advertised price
  compare the returned fields against the fields it promised
  if it falls short, call it again before recording anything
  keep the verdict and the response shape (types only), throw away the values

549 checked: 306 delivered (89% of measurable), 39 short (11%), 204 inconclusive

returned NONE of what they promised (confirmed on two calls):
  stabletravel.dev        /api/flightaware/airports/KMIA/weather   7 promised, 0 back
  api.myceliasignal.com   /oracle/price/btc/usd                    4 promised, 0 back
  vape-x402.vapex402...   /scan/exploit_check                      4 promised, 0 back
  x402.aispace.bot        /api/v1/audio/speech                     3 promised, 0 back
  voice.forgemesh.io      /v1/tts/base                             2 promised, 0 back

sent some, dropped the rest:
  agent.kihustle.tech     missing 14 of 16
  agents.ai-rook.com      missing 3 of 5
  api.locus.report        missing 3 of 4

could not measure (need a key, rejected our input, or 4xx/5xx): 204
these are not counted against the seller
Endpoints measured
345
Delivered everything
306 (89%)
Fell short
39 (11%)
Could not be measured
204
Returned more than promised
178
Every short confirmed on
2 paid calls
Re-verification cost, on-chain
$0.86
Shows up on a free check
none of it
Receipts: 345 endpoints paid and checked field by field against their own published schemas. Every shortfall confirmed on a second paid call. Payments capped at each endpoint's advertised price. Re-verification reconciled on-chain at $0.86.