Evidence registry

AI model evidence coverage

See what kind of evidence is documented for tracked models before relying on a comparison. Official facts, user reports, and controlled tests remain separate.

Evidence registry

Evidence coverage by model

More sources do not mean a model is better. Community reports are not controlled benchmarks, and provider documentation is not independent performance evidence.

ModelOfficial factsCommunity reportsControlled StackLens testLast reviewed
Claude Fable 525Not run2026-07-13
Claude Sonnet 525Not run2026-07-13
Gemini 3.5 Flash25Not run2026-07-13
GPT-5.6 Sol15Not run2026-07-13
GPT-5.6 Terra13Not run2026-07-13
GPT-5.6 Luna13Not run2026-07-13

Read the methodology for source boundaries. A community report is a lead for testing, not a controlled performance result.

Use the result carefully

One dimension is not a complete decision

Cost, token volume, capacity, and evidence coverage answer different questions. Compare them with a representative prompt, actual output length, retry behavior, and your quality bar.

Find a model for a task
FAQ

Ranking methodology questions

Does a higher ranking mean a better model?

No. Each board measures one defined dimension. Token count, price, context capacity, and evidence coverage do not substitute for testing your own completed tasks.