Developer Tools
gremlord
Last updated 2026-09-08 · benchmark measured 2026-09-08 — deterministic & reproducible
Gremlord routes Claude Code requests across multiple AI models with budget controls.
Is gremlord production-ready?
Legit.Show scores gremlord 69 out of 100 — the simple average of its 5 measured frames. Legit.Show ran its deterministic 7-Frame production-readiness benchmark on gremlord (public-surface assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Reliability; its weakest is Privacy. 5 of the seven frames returned a score; Performance and Accessibility were not measurable on this service and are recorded as null — not as zero. Every frame it averages is published with its evidence on the Legit.Show listing.
The 7 Frames
- Performance — not measurable on this service (null — not scored as 0)
- Accessibility — not measurable on this service (null — not scored as 0)
- Security — 45/100
- Privacy — 25/100
- Reliability — 100/100
- Standards — 75/100
- Discoverability — 100/100
What we measured
- Security headers present: HSTS.
- No Content-Security-Policy.
- Served over HTTPS with a valid certificate.
- Real Lighthouse performance run — 695 ms to first byte.
- Returns a proper 404 for unknown routes.
- 1 of 1 sampled routes reachable.
- No privacy policy found.
- Sets cookies / loads scripts with no consent prompt.
Who built it
Maor Bril (@MaorBril)
Who it's for
Claude Code users · Software engineers · Developers using multiple LLM providers · Budget-conscious teams · Anyone needing model flexibility
Visit gremlord → · Alternatives to gremlord → · How this was measured →