# AI Model Rankings by Cost, Context, and Freshness

Compare objective AI model rankings from catalog fields: lowest paid input-token prices, largest recorded context windows, and most recently updated model records.

- AI model ranking
- Records in this view: 637
- Providers shown: 13
- Latest recorded update: 2026-08-03

## How the objective rankings work

The page shows separate rankings for measurable catalog fields and does not create an overall quality score.

### Inclusion criteria

- Each section uses active text-generation offers or grouped model identities.
- A record must have the specific price, context, or date field used by its section.

### Exclusions

- The page excludes subjective quality ranks because the catalog does not contain a common benchmark.
- The context section excludes values above 10 million tokens as unverified outliers.

### Data rules

- **Separate dimensions:** Cost, context, and catalog freshness remain separate. The page does not blend them into one score.
- **Top ten:** Each section shows ten records ordered only by its named catalog field.
- **No quality rank:** A position on this page does not mean that one model produces better answers than another model.

### Limits of the data

- Catalog freshness is not the same as model release date or model quality.
- Price and context do not measure accuracy, speed, safety, or application fit.

## What is ranked

This page ranks recorded fields, not answer quality. Use each section to make a shortlist. Then test the models on your own prompts and constraints.

## Lowest paid input-token prices

Provider offers ordered by the known input-token price.

Showing 10 of 10 records.

| Model | Model ID | Providers | Why listed | Context | Input / output price | Updated |
| --- | --- | --- | --- | ---: | ---: | --- |
| [inclusionAI: Ling-2.6-flash](https://llmcatalog.dev/models/openrouter/inclusionai/ling-2.6-flash) | `inclusionai/ling-2.6-flash` | openrouter | Known paid token prices | 262,144 | $0.01 / $0.03 | 2026-04-21 |
| [Granite 4.0 H Micro](https://llmcatalog.dev/models/cloudflare_workers_ai/@cf/ibm-granite/granite-4.0-h-micro) | `@cf/ibm-granite/granite-4.0-h-micro` | cloudflare_workers_ai | Known paid token prices | 131,000 | $0.02 / $0.11 | 2025-10-07 |
| [IBM: Granite 4.0 Micro](https://llmcatalog.dev/models/openrouter/ibm-granite/granite-4.0-h-micro) | `ibm-granite/granite-4.0-h-micro` | openrouter | Known paid token prices | 131,000 | $0.02 / $0.11 | 2025-10-20 |
| [Mistral: Mistral Nemo](https://llmcatalog.dev/models/openrouter/mistralai/mistral-nemo) | `mistralai/mistral-nemo` | openrouter | Known paid token prices | 131,072 | $0.02 / $0.03 | 2024-07-01 |
| [z-ai/glm-4.6v-flash](https://llmcatalog.dev/models/zenmux/z-ai/glm-4.6v-flash) | `z-ai/glm-4.6v-flash` | zenmux | Known paid token prices | 200,000 | $0.02 / $0.21 | 2025-12-08 |
| [Nex AGI: Nex-N2-Mini](https://llmcatalog.dev/models/openrouter/nex-agi/nex-n2-mini) | `nex-agi/nex-n2-mini` | openrouter | Known paid token prices | 262,144 | $0.03 / $0.10 | 2026-06-24 |
| [Llama 3.2 1B Instruct](https://llmcatalog.dev/models/cloudflare_workers_ai/@cf/meta/llama-3.2-1b-instruct) | `@cf/meta/llama-3.2-1b-instruct` | cloudflare_workers_ai | Known paid token prices | 60,000 | $0.03 / $0.20 | 2024-09-25 |
| [Meta: Llama 3.2 1B Instruct](https://llmcatalog.dev/models/openrouter/meta-llama/llama-3.2-1b-instruct) | `meta-llama/llama-3.2-1b-instruct` | openrouter | Known paid token prices | 60,000 | $0.03 / $0.20 | 2024-09-25 |
| [Llama Prompt Guard 2 22M](https://llmcatalog.dev/models/groq/meta-llama/llama-prompt-guard-2-22m) | `meta-llama/llama-prompt-guard-2-22m` | groq | Known paid token prices | 512 | $0.03 / $0.03 | 2025-05-29 |
| [LFM2-24B-A2B](https://llmcatalog.dev/models/togetherai/LiquidAI/LFM2-24B-A2B) | `LiquidAI/LFM2-24B-A2B` | togetherai | Known paid token prices | 32,768 | $0.03 / $0.12 | 2026-02-25 |

## Largest recorded context windows

Model identities ordered by context. Values above 10 million tokens are excluded.

Showing 10 of 10 records.

| Model | Model ID | Providers | Why listed | Context | Input / output price | Updated |
| --- | --- | --- | --- | ---: | ---: | --- |
| [Grok 4 Fast](https://llmcatalog.dev/models/zenmux/x-ai/grok-4-fast) | `grok-4-fast` | zenmux | Recorded context window | 2,000,000 | $0.20 / $0.50 | 2025-09-19 |
| [Grok 4.1 Fast](https://llmcatalog.dev/models/zenmux/x-ai/grok-4.1-fast) | `grok-4.1-fast` | zenmux | Recorded context window | 2,000,000 | $0.20 / $0.50 | 2025-11-20 |
| [Grok 4.1 Fast Non Reasoning](https://llmcatalog.dev/models/zenmux/x-ai/grok-4.1-fast-non-reasoning) | `grok-4.1-fast-non-reasoning` | zenmux | Recorded context window | 2,000,000 | $0.20 / $0.50 | 2025-11-20 |
| [Grok 4.20](https://llmcatalog.dev/models/venice/grok-4-20) | `grok-4-20` | venice | Recorded context window | 2,000,000 | $1.42 / $2.83 | 2026-06-11 |
| [Grok 4.20 (Reasoning)](https://llmcatalog.dev/models/openrouter/x-ai/grok-4.20) | `grok-4.20` | openrouter, xai | Recorded context window | 2,000,000 | $1.25 / $2.50 | 2026-03-31 |
| [Grok 4.20 Multi-Agent](https://llmcatalog.dev/models/venice/grok-4-20-multi-agent) | `grok-4-20-multi-agent` | venice | Recorded context window | 2,000,000 | $1.42 / $2.83 | 2026-06-11 |
| [Pareto Code Router](https://llmcatalog.dev/models/openrouter/openrouter/pareto-code) | `openrouter/pareto-code` | openrouter | Recorded context window | 2,000,000 | N/A / N/A | 2026-04-21 |
| [SpaceXAI: Grok 4.20 Multi-Agent](https://llmcatalog.dev/models/openrouter/x-ai/grok-4.20-multi-agent) | `grok-4.20-multi-agent` | openrouter, xai | Recorded context window | 2,000,000 | $1.25 / $2.50 | 2026-03-31 |
| [x-ai/grok-4.2-fast](https://llmcatalog.dev/models/zenmux/x-ai/grok-4.2-fast) | `grok-4.2-fast` | zenmux | Recorded context window | 2,000,000 | $3.00 / $9.00 | 2026-03-20 |
| [x-ai/grok-4.2-fast-non-reasoning](https://llmcatalog.dev/models/zenmux/x-ai/grok-4.2-fast-non-reasoning) | `grok-4.2-fast-non-reasoning` | zenmux | Recorded context window | 2,000,000 | $3.00 / $9.00 | 2026-03-20 |

## Most recently updated model records

Model identities ordered by the latest valid catalog date.

Showing 10 of 10 records.

| Model | Model ID | Providers | Why listed | Context | Input / output price | Updated |
| --- | --- | --- | --- | ---: | ---: | --- |
| [Qwen: Qwen3.8 Max](https://llmcatalog.dev/models/alibaba/qwen3.8-max) | `qwen3.8-max` | alibaba, openrouter, zenmux | Latest catalog update | 1,000,000 | $2.00 / $6.00 | 2026-08-03 |
| [Kimi K3 Fast](https://llmcatalog.dev/models/venice/kimi-k3-fast-api) | `kimi-k3-fast-api` | venice | Latest catalog update | 1,000,000 | $4.50 / $22.50 | 2026-08-03 |
| [DeepSeek V4 Flash 0731](https://llmcatalog.dev/models/openrouter/deepseek/deepseek-v4-flash-0731) | `deepseek-v4-flash-0731` | fireworks_ai, openrouter, venice | Latest catalog update | 1,048,576 | $0.09 / $0.18 | 2026-08-01 |
| [deepseek/deepseek-v4-flash-free](https://llmcatalog.dev/models/opencode/deepseek-v4-flash-free) | `deepseek-v4-flash-free` | opencode, zenmux | Latest catalog update | 200,000 | $0.00 / $0.00 | 2026-07-31 |
| [DeepSeek V4 Flash 0731](https://llmcatalog.dev/models/ollama_cloud/deepseek-v4-flash:0731) | `deepseek-v4-flash:0731` | ollama_cloud | Latest catalog update | 1,048,576 | N/A / N/A | 2026-07-31 |
| [DeepSeek V4 Flash 0731](https://llmcatalog.dev/models/togetherai/deepseek-ai/DeepSeek-V4-Flash-0731) | `DeepSeek-V4-Flash-0731` | togetherai | Latest catalog update | 1,000,000 | $0.14 / $0.28 | 2026-07-31 |
| [DeepSeek V4 Flash](https://llmcatalog.dev/models/venice/deepseek-v4-flash) | `deepseek-v4-flash` | deepseek, fireworks_ai, ollama_cloud, opencode, venice, zenmux | Latest catalog update | 1,048,576 | $0.14 / $0.28 | 2026-07-31 |
| [Thinking Machines: Inkling Small](https://llmcatalog.dev/models/openrouter/thinkingmachines/inkling-small) | `inkling-small` | openrouter | Latest catalog update | 524,288 | $0.50 / $1.20 | 2026-07-30 |
| [Kimi K3 Fast](https://llmcatalog.dev/models/fireworks_ai/accounts/fireworks/routers/kimi-k3-fast) | `kimi-k3-fast` | fireworks_ai | Latest catalog update | 1,048,576 | $4.50 / $22.50 | 2026-07-27 |
| [Kimi K3](https://llmcatalog.dev/models/fireworks_ai/accounts/fireworks/models/kimi-k3) | `kimi-k3` | fireworks_ai, moonshotai, ollama_cloud, openrouter, venice, zenmux | Latest catalog update | 1,048,576 | $3.00 / $15.00 | 2026-07-27 |

## Sources

- [llm_db](https://github.com/agentjido/llmdb) — Source for prices, context limits, model dates, providers, and capabilities. Checked 2026-07-30.

See [About LLM Catalog](https://llmcatalog.dev/about) for more information about the data.

## Related LLM model lists

- [Browse the full LLM models list](https://llmcatalog.dev/llm-models) — See active text-generation identities with providers, context, capabilities, and prices.
- [Compare the cheapest LLM APIs](https://llmcatalog.dev/rankings/cheapest-llm-api) — Order paid text-generation offers by known input-token price and check output cost.
- [Compare the largest LLM context windows](https://llmcatalog.dev/models/long-context) — Review active model identities with recorded context limits of 128,000 tokens or more.
- [Find tool-calling LLM models](https://llmcatalog.dev/models/tool-calling) — Compare active text models with explicit tools capability metadata.
- [Browse vision LLM models](https://llmcatalog.dev/models/vision) — Find active text models that accept image input and return text.
- [Browse open-weight LLM models](https://llmcatalog.dev/models/open-weights) — Find active text models whose catalog metadata marks open weights as true.
- [Explore video AI models](https://llmcatalog.dev/models/video) — Use separate lists for models that accept video and models that generate video.

Canonical URL: https://llmcatalog.dev/rankings/ai-models

MCP endpoint: `https://llmcatalog.dev/api/mcp`
