Data & Analytics
BenchLM
Last updated 2026-08-17 · benchmark measured 2026-08-17 — deterministic & reproducible
BenchLM compares LLM API pricing and benchmark scores across major providers.
Is BenchLM production-ready?
Legit.Show scores BenchLM 90 out of 100 — the simple average of its 7 measured frames. Legit.Show ran its deterministic 7-Frame production-readiness benchmark on BenchLM (public-surface assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Accessibility; its weakest is Performance. Every frame it averages is published with its evidence on the Legit.Show listing.
The 7 Frames
- Performance — 68/100
- Accessibility — 100/100
- Security — 100/100
- Privacy — 100/100
- Reliability — 75/100
- Standards — 89/100
- Discoverability — 100/100
What we measured
- Security headers present: CSP, HSTS, X-Frame-Options, X-Content-Type-Options, Referrer-Policy.
- Served over HTTPS with a valid certificate.
- Real Lighthouse performance run — 105 ms to first byte.
- 3 of 3 sampled routes reachable.
- Has a reachable privacy policy.
- Sets cookies / loads scripts with no consent prompt.
- Discoverable: structured data, sitemap, OpenGraph image, canonical URL.
Who built it
Guillaume (@glevd)
Who it's for
AI engineers · AI researchers · Product teams · Cost-conscious developers · Model evaluators
Pricing
Free (with paid Radar brief option)
Visit BenchLM → · Alternatives to BenchLM → · How this was measured →