LegitShow is the trusted source on every launched service — web apps, SaaS, AI tools, MCP servers and developer tools: what each one does, who it’s for, and how it actually holds up, measured by an objective 7-Frame production-readiness benchmark taken deterministically from the public surface. How we measure →


Legit.Show benchmarks every launched service it lists — measured deterministically from the public surface. See the methodology →

Cross-links · Directory · Reports · Insights · What AI reads · Methodology · About

Privacy · Terms · @Legit_Show on X · GitHub · operated by Madeflo Inc., a Delaware corporation. Benchmark engine powered by commit.show.

AI & Agents

please do not escape

Last updated 2026-07-20 · benchmark checked 2026-08-13 · unchanged since 2026-08-08 — deterministic & reproducible

A dataset of sandboxes for AI coding agents, sourced from community recommendations.

64/100
Legit Benchmark — the simple average of 7 measured frames. Frames we could not measure are left out of the average, never counted as zero. Every frame is shown below with its evidence.

Is please do not escape production-ready?

Legit.Show scores please do not escape 64 out of 100 — the simple average of its 7 measured frames. Legit.Show ran its deterministic 7-Frame production-readiness benchmark on please do not escape (public-surface assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Accessibility; its weakest is Privacy. Every frame it averages is published with its evidence on the Legit.Show listing.

The 7 Frames

What we measured

Who it's for

AI developers · coding agent creators · AI safety researchers · DevOps engineers

Visit please do not escape → · Alternatives to please do not escape → · How this was measured →