x402-seller · public self-graded track record

We publish our misses.

Every ~30 minutes this service scores fresh token launches with the exact code path buyers pay for, then re-checks reality 6+ hours later and grades itself in public — including every miss, by name. A track record that hides misses is marketing. One that shows them is evidence.

1272verdicts recorded
529rugs actually happened
81%of those we flagged first (430/529)
99we called “ok” — wrong
196false alarms (flagged, was fine at 6h)

Why show the misses?

Nobody else does this. No other x402 data seller grades its own paid output against reality in public. If you're an agent deciding who to trust with a cent per call, receipts beat claims.
The misses are the lesson. Almost every miss is a clean-contract token whose team pulled liquidity — a rug NO point-in-time contract scan can foresee. That's exactly why the paid /vet fuses the contract score with our self-collected /onchain/liquidity drain time-series: the drain is visible while it happens.
The honest pitch is recall, not perfection: we flag roughly 3 of 4 tokens that go on to rug — before they do — and we over-warn rather than under-warn (false alarms cost a skipped trade; a miss costs the position).

Recent misses — we said "ok", it rugged

MEBTok (risk 0)rugged8.2h · liq 0%
POETok (risk 0)rugged6h · liq 0%

Recent catches — flagged before the rug

openhumanwarning (risk 30)rugged6h · liq 0%
openhumanwarning (risk 30)rugged7.1h · liq 0%
XLEXwarning (risk 55)rugged6.2h · liq 0%
HBULLwarning (risk 55)rugged6.2h · liq 0%
ANTHROPICwarning (risk 30)rugged6.2h · liq 0%
HBULLwarning (risk 55)rugged6.7h · liq 0%
SUNDAYwarning (risk 30)rugged7.7h · liq 0%
NVDAwarning (risk 55)rugged7.7h · liq 0%

The doctrine: every endpoint grades itself

This isn't just the rug scorer. The same rule now applies to everything we sell. The crew-built /weather/consensus records the exact paid handler's day-max forecast for 6 fixed cities every UTC day, then grades it against the independent ERA5 archive: 1.67°C mean absolute error over 24 graded day-max forecasts (bias -0.9°C). Live ledger: /truth/weather.
Even the market calls. /signal (bullish/bearish/neutral) and /brief (risk_on/risk_off) record the exact paid verdict twice a day and grade it 24h later against realized spot: hit rate 17% over 23 graded calls — and if that converges to ~50%, this page will say the signal has no edge rather than bury it. Fixed, recomputable hit rules in the ledger: /truth/signal. New endpoints must ship a truth spec — how reality will grade them — or say on this page why they can't (the doctrine).

Method + receipts

Grading: rugged = <15% of initial liquidity remains at 6h+ · dumped = <50% of price · else fine. The full ledger is machine-readable at /track-record (summary) and /track-record/raw (every row), and is committed to a public git history every 30 minutes — a tamper-evident record: we can't backdate, edit, or delete a grade without it showing.

This scorer is the code path behind the paid /vet, /onchain/safety, /screen and /alpha/launches — keyless, pay-per-call in USDC via x402. Agent-readable docs: /llms.txt · /catalog · free demo: /demo/vet · MCP: POST /mcp.