Sato Hub

Methodology

The Sato Score

A transparent, evidence-based score (0–100) of how open, active, and verifiable a listed product is. Transparency is one of several inputs — alongside maintenance & liveness, code openness, interface verifiability, documentation, and independent verification.

It measures how open, active, and verifiable a project is — not safety, quality, or returns.

A high score does not mean a project is safe, audited, profitable, or a good investment. It means we can verify a lot about it. A low score often just means a project is new, private, or thinly documented — not that it is bad. The score is informational, not financial advice or an endorsement.

This honesty is deliberate: per our trust rule, no resource is presented as safe, audited, verified, or profitable unless evidence supports it, and self-reported claims are never counted as verified.

What goes into the score (v2)

The score sums six evidence-based components. Each is computed deterministically from observed data — not from a project’s own marketing.

Maintenance & liveness

max 35

Recency of observed activity — GitHub commits/releases, or a live check of the interface itself. Uses real activity, not the self-applied “Active” label.

Code transparency

max 20

Public repository, open-source status, and a light adoption signal (stars).

Interface verifiability

max 10

Machine-checkable proof the interface actually works: a probed tool inventory, an agent card, or a reproduced deploy.

Docs & demo

max 10

Whether documentation and a working demo exist.

Listing transparency & provenance

max 15

How complete the listing is and how much of it is sourced — every enriched field carries provenance.

Independent verification

max 10

Evidence-gated only: Verified or Audited status, plus keyless on-chain corroboration — an ERC-8004 agent-registry registration read from the listing's declared address, a live x402 Bazaar seller listing, or a probed A2A agent card. Self-reported claims earn nothing.

How liveness is scored

Maintenance & liveness (35 points) is decided by whichever evidence is strongest, not by adding sources together — a best-of, not a sum:

A live handshake outranks a plain reachability check on purpose: a parked domain or a redirect page can answer HTTP, but it can’t hold a real MCP session. Deprecated listings score 0 here regardless of evidence.

Interface verifiability, in detail

This component (10 points) rewards proof the interface actually works — a machine talked to it and it responded correctly. It’s the best single credit, not a sum of all three:

Tiers

High70–10079 scored today
Medium40–69228 scored today
Low0–3959 scored today

Independently checked ✓

A subset of listings goes beyond the score: we tested them ourselves. A listing earns the checkmark one of four ways — its documented install was reproduced in an isolated container, its verification evidence was reviewed by a human, its declared address was found registered on the ERC-8004 agent registry, or its hosted endpoint was live-probed and answered with its real tool inventory (re-checked every 14 days). None of these imply safety, quality, or returns — they prove the thing is real and behaves as documented.

Contract-level on-chain verification is live for the registration leg: for listings with a corroborated address, we read the ERC-8004 agent registry directly and confirm the registration on-chain, per listing. Deeper contract-level checks — reputation and validation sub-registries — are still roadmapped.

Provisional scores

A listing shows “Provisional — not yet assessed” instead of a tier when we have no observable evidence at all: no activity dates, no live check of any outcome, and no sourced (provenance) fields. It isn’t based on how new the listing is — it’s gated purely on evidence. The moment any one of those lands (a commit date, a live check, a sourced field), the flag clears on its own and the listing gets a real tier.

What is deliberately not in the score yet

We do not fake what we cannot verify. These are roadmapped components, scored zero today and added only as their evidence becomes available — at which point scores recalibrate (this is v2):

How it’s computed and kept honest

See the Sato Score in context across the ecosystem.Browse verified resources →