Audit ledger · Generated from the live database · September 29, 2026

What reviewing 2,079 Claude tools actually looks like.

The Claude Observatory tracks 3,482 tools — skills, MCP servers, hooks, patterns, and workflows. 2,079 of them have reached a verdict. This page is the ledger: every number below is queried from the review database the moment the site builds, so it can't go stale and it can't be massaged.


1,684 Approved

Published with a grade, a review depth, and caveats where they're due.

365 Rejected

Didn't clear the bar. They stay in the database — the no's are part of the record.

30 Security holds

Frozen pending a security question we couldn't resolve. Not listed until it is.

01The full ledger

Every tool, every status

1,383 tools are still in the pipeline — evaluating or staged for triage. They don't get listed until they get a verdict.

  • Approved · 1,684 · 48.4%
  • Rejected · 365 · 10.5%
  • Security hold · 30 · 0.9%
  • Evaluating · 1,383 · 39.7%
  • Deprecated · 20 · 0.6%

Of the 2,079 tools that reached a verdict, 19.0% didn't make it — roughly one in 5. A catalog that approves everything isn't a review; it's a directory.

02Grades of the approved

Approval is not endorsement

Each approved tool carries a letter grade from its latest evaluation — a weighted score across client readiness, breadth of use, reliability, and security.

Grade A Recommended without hesitation 613.6%
Grade B Solid — minor caveats 45827.2%
Grade C Usable — know the limits 1,00459.6%
Grade D Approved with warnings attached 1519.0%
Grade F Kept only for the record 100.6%

Grade C is the biggest bucket at 59.6% of approvals. Only 61 tools have earned an A. That's the honest shape of this ecosystem right now: mostly usable, rarely exceptional.

03Review depth

How closely each tool was examined

Not every review is the same review, and pretending otherwise would be dishonest. Every listing on the site discloses its depth.

Tested Installed and run hands-on, end to end 130.8%
Reviewed Source and docs read closely, not run 17810.6%
Scanned Automated signals + structured pass 1,48488.1%
Listed Catalogued with basic metadata only 90.5%

Only 13 tools have been personally tested end to end; 88.1% are scanned. Hands-on testing is the scarcest resource in this catalog — which is exactly why we label it.

04Domain skew

Where the catalog leans

Approved tools across the eight domains we track. The skew is real, so we show it.

Code & development 87351.8%
Productivity & workflow 21412.7%
Data & analytics 1458.6%
Documents & content 1217.2%
Security & compliance 1016.0%
Infrastructure & DevOps 925.5%
Config & setup 814.8%
Communication & collaboration 573.4%

Code & development alone is 51.8% of everything approved. The Claude tool ecosystem still builds mostly for developers — a gap worth knowing about if you're shopping for anything else.

05Trust signals

What the factual record shows

Alongside reviews, the pipeline collects factual signals from GitHub and npm for 1,684 of the 1,684 approved tools. Three that matter:

85.0% Known license

1,432 approved tools have an identifiable license. The rest — you're deploying on trust.

87.7% Commit in last 90 days

1,477 approved tools show recent maintainer activity as of the latest scan.

0 Known CVEs recorded

Across all approved tools' registry records at last scan. Absence of a CVE is not a security guarantee — see review depth above.

06The work behind it

The paper trail

Evaluations logged Append-only scoring history — re-reviews included 5,955
Field notes Tried it; here's what happened 179

Trust-signal scans for approved tools run continuously: the oldest current scan dates to September 21, 2026, the newest to September 29, 2026. Tools that fail on re-review get downgraded or moved to the not-recommended list — the grade you see is the latest, not the best.

How to cite this

This page is built to be referenced. Cite it as:

Matthews, A. "The Dataset Report: What reviewing 2,079 Claude tools actually looks like." Value Alignment Consulting, September 29, 2026. https://valuealignmentconsulting.com/dataset-report

Numbers regenerate from the live review database on every site build, so figures on this page move as the catalog grows. The methodology is documented on the evaluation guide and about pages.

Rolling Claude out in your org? Let's talk.

Start a conversation →