Legit.Show is a directory of launched web apps, SaaS, AI tools, MCP servers and developer tools — each with an objective 7-Frame production-readiness benchmark, measured deterministically from the public surface. How we measure →


Legit.Show benchmarks every launched service it lists — measured deterministically from the public surface. See the methodology →

Cross-links · Directory · Reports · Methodology · About

Privacy · Terms · operated by Madeflo Inc., a Delaware corporation. Benchmark engine powered by commit.show.

AI & Agents

Galileo AI

Last updated 2026-08-03 · benchmark measured 2026-08-03 — deterministic & reproducible

AI observability and evaluation platform for monitoring and improving LLM applications.

73/100
Legit Benchmark — the simple average of 7 measured frames. Frames we could not measure are left out of the average, never counted as zero. Every frame is shown below with its evidence.

Is Galileo AI production-ready?

Legit.Show scores Galileo AI 73 out of 100 — the simple average of its 7 measured frames. Legit.Show ran its deterministic 7-Frame production-readiness benchmark on Galileo AI (public-surface assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Discoverability; its weakest is Performance. Every frame it averages is published with its evidence on the Legit.Show listing.

The 7 Frames

What we measured

Who it's for

Developers · Small teams · Enterprise teams · AI teams · ML engineers

Pricing

Free ($0/month for 5,000 traces), Pro ($100/month for 50,000 traces), Enterprise (custom)

Visit Galileo AI → · How this was measured →