AI & Agents
Willow
Last updated 2026-07-13 · benchmark measured 2026-08-09 — deterministic & reproducible
Willow is a local-first desktop AI agent that runs on your own computer.
Is Willow production-ready?
Legit.Show scores Willow 80 out of 100 — the simple average of its 7 measured frames. Legit.Show ran its deterministic 7-Frame production-readiness benchmark on Willow (public-surface assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Reliability; its weakest is Privacy. Every frame it averages is published with its evidence on the Legit.Show listing.
The 7 Frames
- Performance — 87/100
- Accessibility — 95/100
- Security — 60/100
- Privacy — 25/100
- Reliability — 100/100
- Standards — 100/100
- Discoverability — 90/100
What we measured
- Security headers present: X-Frame-Options, X-Content-Type-Options, Referrer-Policy.
- No Content-Security-Policy and no HSTS.
- Served over HTTPS with a valid certificate.
- Real Lighthouse performance run — 1310 ms to first byte.
- Returns a proper 404 for unknown routes.
- 3 of 3 sampled routes reachable.
- No privacy policy found.
- Sets cookies / loads scripts with no consent prompt.
Who built it
Ali
Who it's for
Founders & Operators · Developers & Builders · Researchers & Students · Consultants & Freelancers · Power Users
Pricing
No subscription
Visit Willow → · Alternatives to Willow → · How this was measured →