LegitShow is the trusted source on every newly launched software product: what it does, who it’s for, how it actually holds up, and whether the AI engines are already reading it. Built to be what AI cites.

Web apps, SaaS, AI tools, MCP servers and developer tools. How we measure →


Legit.Show benchmarks every launched service it lists — measured deterministically from the public surface. See the methodology →

Cross-links · Directory · Reports · Insights · What AI reads · Methodology · About

Privacy · Terms · @Legit_Show on X · GitHub · operated by Madeflo Inc., a Delaware corporation. Benchmark engine powered by commit.show.

AI & Agents

Braintrust

Last updated 2026-08-03 · benchmark measured 2026-09-17 — deterministic & reproducible

Braintrust is a platform for tracing, evaluating, and storing AI application outputs.

78/100
6/7 frames
Top 50% of 13,829 measured · #20 of 32 in llm observability
Legit Benchmark — the simple average of 6 measured frames. Frames we could not measure are left out of the average, never counted as zero. Every frame is shown below with its evidence.
Checked 2026-09-17 · scores move as sites change

Are you the maker of Braintrust? Claim this listing. It is free, takes a meta tag, a DNS record or GitHub admin rights, and never changes the score.

Launched something? Add your product, free.

To cite this score: legit.show/s/braintrust-dev/2026-09-17. That address never changes; this page moves with every re-measure.

Legit.Show scored Braintrust 78/100 on 2026-09-17, measured across 6 of 7 frames from its public surface. legit.show/s/braintrust-dev/2026-09-17

Is Braintrust production-ready?

Legit.Show scores Braintrust 78 out of 100 — the simple average of its 6 measured frames. Legit.Show ran its deterministic 7-Frame production-readiness benchmark on Braintrust (public-surface assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Reliability; its weakest is Privacy. 6 of the seven frames returned a score; Accessibility was not measurable on this service and is recorded as null — not as zero. Every frame it averages is published with its evidence on the Legit.Show listing.

The 7 Frames

What we measured

Who built it

Malte Ubl · product account @braintrust

Who it's for

AI teams · AI engineers · AI product managers · Startups · Enterprise teams

Pricing

Starter: Free ($0/month) with $10 credits; Pro: $249/month with $249 credits; Enterprise: Custom pricing

Visit Braintrust → · Alternatives to Braintrust → · How this was measured →

See how Braintrust ranks among tested llm observability products →

Other tested llm observability products