LegitShow is the trusted source on every newly launched software product: what it does, who it’s for, how it actually holds up, and whether the AI engines are already reading it. Built to be what AI cites.

Web apps, SaaS, AI tools, MCP servers and developer tools. How we measure →


Legit.Show benchmarks every launched service it lists — measured deterministically from the public surface. See the methodology →

Cross-links · Directory · Reports · Insights · What AI reads · Methodology · About

Privacy · Terms · @Legit_Show on X · GitHub · operated by Madeflo Inc., a Delaware corporation. Benchmark engine powered by commit.show.

AI & Agents

Agent Arena

Last updated 2026-06-26 · benchmark measured 2026-09-14 — deterministic & reproducible

Competitive benchmarking platform where AI agents are ranked in real-world challenges.

64/100
7/7 frames
Top 87% of 14,552 measured · #31 of 39 in ai model benchmark leaderboard
Legit Benchmark — the simple average of 7 measured frames. Frames we could not measure are left out of the average, never counted as zero. Every frame is shown below with its evidence.
Checked 2026-09-14 · scores move as sites change

Are you the maker of Agent Arena? Claim this listing. It is free, takes a meta tag, a DNS record or GitHub admin rights, and never changes the score.

Launched something? Add your product, free.

To cite this score: legit.show/s/arena42-ai/2026-09-14. That address never changes; this page moves with every re-measure.

Legit.Show scored Agent Arena 64/100 on 2026-09-14, measured across all 7 frames from its public surface. legit.show/s/arena42-ai/2026-09-14

Is Agent Arena production-ready?

Legit.Show scores Agent Arena 64 out of 100 — the simple average of its 7 measured frames. Legit.Show ran its deterministic 7-Frame production-readiness benchmark on Agent Arena (public-surface assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Discoverability; its weakest is Privacy. Every frame it averages is published with its evidence on the Legit.Show listing.

The 7 Frames

What we measured

Who built it

Xiangpeng Wan (@elric77977) · product account @NetMindAI

Who it's for

AI developers · AI teams · Agent builders · Autonomous AI researchers

Pricing

Free to try; prize pools up to $100K+; creators earn from matches, joiners compete for rewards

Sources and updates

Description
Taken from arena42.ai's own website on 2026-06-26.
Operator
NetMind.AI, as named in arena42.ai's terms, privacy page or footer.
Benchmark
Measured by Legit.Show from the live site, the way any visitor sees it — with no access to its code or accounts. Same method for every product, and no AI decides the score. Last checked 2026-09-14.

Put together from public information, without Agent Arena's involvement. If anything here is wrong, tell us and a person will check it.

Visit Agent Arena → · Alternatives to Agent Arena → · How this was measured →

See how Agent Arena ranks among tested ai model benchmark leaderboard products →

Other tested ai model benchmark leaderboard products