AI & Agents
caveman
Last updated 2026-06-05 · benchmark measured 2026-06-10 — deterministic & reproducible
Token-efficient tooling stack for AI agent development that reduces LLM costs.
Overall production-readiness score — reserved
Legit.Show benchmarked caveman across all seven frames. The single overall score is shown to the verified maker; the frame-by-frame breakdown is public below.
Is caveman production-ready?
Legit.Show ran its deterministic 7-Frame production-readiness benchmark on caveman (public-surface assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Performance; its weakest is Privacy. The per-frame breakdown is public on the Legit.Show listing; the single overall production-readiness score is reserved for the verified maker.
The 7 Frames
- Performance — 93/100
- Accessibility — 93/100
- Security — 45/100
- Privacy — 0/100
- Reliability — 92/100
- Standards — 92/100
- Discoverability — 15/100
What we measured
- Security headers present: HSTS.
- No Content-Security-Policy.
- Served over HTTPS with a valid certificate.
- Real Lighthouse performance run — 64 ms to first byte.
- Returns a proper 404 for unknown routes.
- 0 of 0 sampled routes reachable.
- No privacy policy found.
- Sets cookies / loads scripts with no consent prompt.
Who it's for
AI engineers · Agent developers · LLM practitioners · Cost-conscious teams · DevOps engineers
Pricing
Free installation; pay-as-you-use metered token pricing; $0 until live