Audit ledger · Generated from the live database · August 30, 2026

What reviewing 1,893 Claude tools actually looks like.

The Claude Observatory tracks 3,032 tools — skills, MCP servers, hooks, patterns, and workflows. 1,893 of them have reached a verdict. This page is the ledger: every number below is queried from the review database the moment the site builds, so it can't go stale and it can't be massaged.


1,512 Approved

Published with a grade, a review depth, and caveats where they're due.

352 Rejected

Didn't clear the bar. They stay in the database — the no's are part of the record.

29 Security holds

Frozen pending a security question we couldn't resolve. Not listed until it is.

01The full ledger

Every tool, every status

1,119 tools are still in the pipeline — evaluating or staged for triage. They don't get listed until they get a verdict.

  • Approved · 1,512 · 49.9%
  • Rejected · 352 · 11.6%
  • Security hold · 29 · 1.0%
  • Evaluating · 1,119 · 36.9%
  • Deprecated · 20 · 0.7%

Of the 1,893 tools that reached a verdict, 20.1% didn't make it — roughly one in 5. A catalog that approves everything isn't a review; it's a directory.

02Grades of the approved

Approval is not endorsement

Each approved tool carries a letter grade from its latest evaluation — a weighted score across client readiness, breadth of use, reliability, and security.

Grade A Recommended without hesitation 553.6%
Grade B Solid — minor caveats 39826.3%
Grade C Usable — know the limits 90960.1%
Grade D Approved with warnings attached 1409.3%
Grade F Kept only for the record 100.7%

Grade C is the biggest bucket at 60.1% of approvals. Only 55 tools have earned an A. That's the honest shape of this ecosystem right now: mostly usable, rarely exceptional.

03Review depth

How closely each tool was examined

Not every review is the same review, and pretending otherwise would be dishonest. Every listing on the site discloses its depth.

Tested Installed and run hands-on, end to end 130.9%
Reviewed Source and docs read closely, not run 17811.8%
Scanned Automated signals + structured pass 1,31286.8%
Listed Catalogued with basic metadata only 90.6%

Only 13 tools have been personally tested end to end; 86.8% are scanned. Hands-on testing is the scarcest resource in this catalog — which is exactly why we label it.

04Domain skew

Where the catalog leans

Approved tools across the eight domains we track. The skew is real, so we show it.

Code & development 77451.2%
Productivity & workflow 19512.9%
Data & analytics 1409.3%
Documents & content 1056.9%
Security & compliance 906.0%
Infrastructure & DevOps 835.5%
Config & setup 714.7%
Communication & collaboration 543.6%

Code & development alone is 51.2% of everything approved. The Claude tool ecosystem still builds mostly for developers — a gap worth knowing about if you're shopping for anything else.

05Trust signals

What the factual record shows

Alongside reviews, the pipeline collects factual signals from GitHub and npm for 1,512 of the 1,512 approved tools. Three that matter:

84.2% Known license

1,273 approved tools have an identifiable license. The rest — you're deploying on trust.

88.8% Commit in last 90 days

1,342 approved tools show recent maintainer activity as of the latest scan.

0 Known CVEs recorded

Across all approved tools' registry records at last scan. Absence of a CVE is not a security guarantee — see review depth above.

06The work behind it

The paper trail

Evaluations logged Append-only scoring history — re-reviews included 5,343
Field notes Tried it; here's what happened 179

Trust-signal scans for approved tools run continuously: the oldest current scan dates to August 22, 2026, the newest to August 30, 2026. Tools that fail on re-review get downgraded or moved to the not-recommended list — the grade you see is the latest, not the best.

How to cite this

This page is built to be referenced. Cite it as:

Matthews, A. "The Dataset Report: What reviewing 1,893 Claude tools actually looks like." Value Alignment Consulting, August 30, 2026. https://valuealignmentconsulting.com/dataset-report

Numbers regenerate from the live review database on every site build, so figures on this page move as the catalog grows. The methodology is documented on the evaluation guide and about pages.

Rolling Claude out in your org? Let's talk.

Start a conversation →