Web
FlexInference
Last updated 2026-07-07 · benchmark checked 2026-08-12 · unchanged since 2026-08-05 — deterministic & reproducible
Deadline-aware, OpenAI-compatible LLM router. Bring your own OpenAI, Gemini, or Anthropic key and race a flex tier within your deadline for lower cost.
Is FlexInference production-ready?
Legit.Show scores FlexInference 88 out of 100 — the simple average of its 6 measured frames. Legit.Show ran its deterministic 7-Frame production-readiness benchmark on FlexInference (public-surface assessment), measured from the public surface with no LLM in the scoring path. Its strongest frame is Privacy; its weakest is Security. 6 of the seven frames returned a score; Accessibility was not measurable on this service and is recorded as null — not as zero. Every frame it averages is published with its evidence on the Legit.Show listing.
The 7 Frames
- Performance — 80/100
- Accessibility — not measurable on this service (null — not scored as 0)
- Security — 45/100
- Privacy — 100/100
- Reliability — 100/100
- Standards — 100/100
- Discoverability — 100/100
What we measured
- Security headers present: HSTS.
- No Content-Security-Policy.
- Served over HTTPS with a valid certificate.
- Real Lighthouse performance run — 538 ms to first byte.
- Returns a proper 404 for unknown routes.
- 3 of 3 sampled routes reachable.
- Has a reachable privacy policy.
- Sets cookies / loads scripts with no consent prompt.
Who built it
product account @FlexInference
Pricing
Free (bring your own keys); Flex tier pricing starts at $0.50 per 1M tokens (50% discount if completed within time budget)