Legit.Show is a directory of launched web apps, SaaS, AI tools, MCP servers and developer tools — each with an objective 7-Frame production-readiness benchmark, measured deterministically from the public surface. How we measure →


Legit.Show benchmarks every launched service it lists — measured deterministically from the public surface. See the methodology →

Cross-links · Directory · Reports · Methodology · About

Privacy · Terms · operated by Madeflo Inc., a Delaware corporation. Benchmark engine powered by commit.show.

AI & Agents

claude-code-production-grade-plugin

Last updated 2026-06-15 · benchmark measured 2026-07-31 — deterministic & reproducible

Claude Code Plugin: Fully autonomous production-grade SaaS pipeline — 14 bundled skills, CEO/CTO command-driven, single install

38/100
Legit Benchmark — the simple average of 3 measured frames. Frames we could not measure are left out of the average, never counted as zero. Every frame is shown below with its evidence.

Is claude-code-production-grade-plugin production-ready?

Legit.Show scores claude-code-production-grade-plugin 38 out of 100 — the simple average of its 3 measured frames. Legit.Show ran its deterministic 7-Frame production-readiness benchmark on claude-code-production-grade-plugin (github assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Security; its weakest is Discoverability. 3 of the seven frames returned a score; Performance, Accessibility, Privacy and Reliability were not measurable on this service and are recorded as null — not as zero. Maintenance is an additional frame from the open-source teardown, scored separately from the seven. Every frame it averages is published with its evidence on the Legit.Show listing.

The 7 Frames

Open-source teardown

Scored separately — not one of the seven.

What we measured

Visit claude-code-production-grade-plugin → · How this was measured →