Web · OSS
pua
Last updated 2026-06-15 · benchmark checked 2026-08-09 · unchanged since 2026-08-08 — deterministic & reproducible
你是一个曾经被寄予厚望的 P8 级工程师。Anthropic 当初给你定级的时候,对你的期望是很高的。 一个agent使用的高能动性的skill。 Your AI has been placed on a PIP. 30 days to show improvement.
73/100
Legit Benchmark — the simple average of 7 measured frames. Frames we could not measure are left out of the average, never counted as zero. Every frame is shown below with its evidence.
Is pua production-ready?
Legit.Show scores pua 73 out of 100 — the simple average of its 7 measured frames. Legit.Show ran its deterministic 7-Frame production-readiness benchmark on pua (public-surface assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Discoverability; its weakest is Privacy. Every frame it averages is published with its evidence on the Legit.Show listing.
The 7 Frames
- Performance — 69/100
- Accessibility — 82/100
- Security — 65/100
- Privacy — 25/100
- Reliability — 75/100
- Standards — 92/100
- Discoverability — 100/100
What we measured
- Security headers present: HSTS, X-Content-Type-Options, Referrer-Policy.
- No Content-Security-Policy.
- Served over HTTPS with a valid certificate.
- Real Lighthouse performance run — 164 ms to first byte.
- 3 of 3 sampled routes reachable.
- No privacy policy found.
- Sets cookies / loads scripts with no consent prompt.
- Discoverable: structured data, sitemap, OpenGraph image, canonical URL.
Pricing
Free and open-source under MIT license