Model evidence
Good enough for which task?
Explore measured models and available catalog listings. Quality comes from published task evidence; models without task results are clearly marked.
| Model | Evidence | Cheapest passing | Median quality | Providers |
|---|---|---|---|---|
Model evidence
Explore measured models and available catalog listings. Quality comes from published task evidence; models without task results are clearly marked.
| Model | Evidence | Cheapest passing | Median quality | Providers |
|---|---|---|---|---|
| Model | Evidence | Cheapest passing | Median quality | Providers |
|---|---|---|---|---|
| lightning-ai/qwen3.8-27b Unclassified | Not yet measured | — | — | TrustedRouter |
| lightricks/ltx-2.3 Unclassified | Not yet measured | — | — | TrustedRouter |
| lightricks/ltx-2.3-fast Unclassified | Not yet measured | — | — | TrustedRouter |
| liquid/lfm-2-24b-a2b Unclassified | Not yet measured | — | — | OpenRouter |
| liquid/lfm-2.5-1.2b-instruct:free Unclassified | Not yet measured | — | — | OpenRouter |
| liquid/lfm-2.5-1.2b-thinking:free Unclassified | Not yet measured | — | — | OpenRouter |
| liquid/lfm-2.5-2.6b:free Unclassified | Not yet measured | — | — | OpenRouter |
| llama3-3-70b Unclassified | Not yet measured | — | — | Neurometric |
| llama4-maverick-17b Unclassified | Not yet measured | — | — | Neurometric |
| llama4-scout-17b Unclassified | Not yet measured | — | — | Neurometric |
| lowes-gemma-4-12b-it Unclassified | Not yet measured | — | — | Price unavailable |
| magistral-small-2509 Unclassified | Not yet measured | — | — | Neurometric |
| mancer/weaver Unclassified | Not yet measured | — | — | OpenRouter |
| meituan-longcat/longcat-2.0 Unclassified | Not yet measured | — | — | TrustedRouter |
| meituan/longcat-2.0 Unclassified | Not yet measured | — | — | OpenRouter |
| meta-llama/llama-3-70b-instruct Unclassified | Not yet measured | — | — | OpenRouter |
| meta-llama/llama-3-8b-instruct Unclassified | Not yet measured | — | — | OpenRouterTrustedRouter |
| meta-llama/llama-3.1-70b-instruct Unclassified | Not yet measured | — | — | OpenRouterTrustedRouter |
| meta-llama/llama-3.1-8b-instruct Unclassified | Not yet measured | — | — | OpenRouterTrustedRouter |
| meta-llama/llama-3.1-8b-instruct-fp8 Unclassified | Not yet measured | — | — | TrustedRouter |
Cheapest passing counts tasks where the model clears the task’s published quality floor (80% when none is recorded) and has the lowest observed workload cost. Median quality summarizes those published task results; it is not a general model ranking.