AI & Agents
pi
Last updated 2026-06-06 · benchmark measured 2026-08-09 — deterministic & reproducible
AI agent toolkit: coding agent CLI, unified LLM API, TUI & web UI libraries, Slack bot, vLLM pods
Is pi production-ready?
Legit.Show scores pi 64 out of 100 — the simple average of its 4 measured frames. Legit.Show ran its deterministic 7-Frame production-readiness benchmark on pi (github assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Security; its weakest is Discoverability. 4 of the seven frames returned a score; Performance, Accessibility and Privacy were not measurable on this service and are recorded as null — not as zero. Maintenance is an additional frame from the open-source teardown, scored separately from the seven. Every frame it averages is published with its evidence on the Legit.Show listing.
The 7 Frames
- Performance — not measurable on this service (null — not scored as 0)
- Accessibility — not measurable on this service (null — not scored as 0)
- Security — 80/100
- Privacy — not measurable on this service (null — not scored as 0)
- Reliability — 65/100
- Standards — 75/100
- Discoverability — 35/100
Open-source teardown
Scored separately — not one of the seven.
- Maintenance — 100/100
What we measured
- No Content-Security-Policy and no HSTS.
- No privacy policy found.
- Sets cookies / loads scripts with no consent prompt.