AI & Agents
Raven
Last updated 2026-08-03 · benchmark measured 2026-08-03 — deterministic & reproducible
Raven is a self-improving AI agent harness built on the EverOS memory platform.
71/100
Legit Benchmark — the simple average of 7 measured frames. Frames we could not measure are left out of the average, never counted as zero. Every frame is shown below with its evidence.
Is Raven production-ready?
Legit.Show scores Raven 71 out of 100 — the simple average of its 7 measured frames. Legit.Show ran its deterministic 7-Frame production-readiness benchmark on Raven (public-surface assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Reliability; its weakest is Privacy. Every frame it averages is published with its evidence on the Legit.Show listing.
The 7 Frames
- Performance — 72/100
- Accessibility — 87/100
- Security — 55/100
- Privacy — 25/100
- Reliability — 100/100
- Standards — 89/100
- Discoverability — 72/100
What we measured
- Security headers present: HSTS, X-Content-Type-Options.
- No Content-Security-Policy.
- Served over HTTPS with a valid certificate.
- Real Lighthouse performance run — 141 ms to first byte.
- Returns a proper 404 for unknown routes.
- 1 of 1 sampled routes reachable.
- No privacy policy found.
- Sets cookies / loads scripts with no consent prompt.
Who built it
product account @evermind
Who it's for
Individual users · Enterprise teams · Agent developers · Workflow automation teams · Solo operators