Models tracked

396

Providers

56

New this week

5

Biggest context

2,000K

Free models

21

The market on one plane — context vs price, by what a model can be fed
Text onlyImageDocument + imageAudio / videofilled = open weights · hollow = closed

Both axes are logarithmic. 26 models are off-chart: the free tier and the routers OpenRouter prices at −1, neither of which a log scale admits.

Newest models
Most recently added to the OpenRouter catalog
Model Provider Context $/1M in $/1M out Caps Added
Hy4 preview Tencent 1,049K $0.83 $2.50 🧠 🔧 2026-08-28
Ling 3.0 Flash Fin (free) Inclusionai 262K free free 🧠 🔧 2026-08-27
Qwen3.8 Flash Qwen 1,000K $0.15 $0.47 🧠 🖼️ 🔧 2026-08-26
GLM 5.3 Flash Z.ai 1,311K $0.07 $0.25 🧠 🖼️ 🔧 2026-08-26
GLM 5.3 Flash (batch) Z.ai 1,049K $0.15 $0.50 🧠 🖼️ 🔧 2026-08-26
Muse Spark 1.2 Contributor Meta 1,049K $0.10 $0.20 🧠 🖼️ 🔧 2026-08-21
DeepSeek V4 Flash Vision Exp DeepSeek 1,049K $0.22 $0.66 🧠 🖼️ 🔧 2026-08-21
Hy-MT2-1.8B Tencent 8K $0.04 $0.18 2026-08-20
Hy-MT2-30B-A3B Tencent 8K $0.07 $0.29 2026-08-20
GLM Latest Z.ai 1,311K $1.19 $4.18 🧠 🔧 2026-08-19
Hy-MT2-7B Tencent 8K $0.07 $0.29 2026-08-19
GLM 5.3 Z.ai 1,311K $1.40 $4.40 🧠 🔧 2026-08-18
🧠 reasoning · 🖼️ multimodal · 🔧 tools · Data: OpenRouter
Cheapest capable models
≥100K context, lowest input price ($/1M tokens)
Model Provider Context $/1M in $/1M out Caps Added
Granite 4.0 Micro IBM 131K $0.02 $0.11 2025-10-20
Mistral Nemo Mistral 131K $0.02 $0.03 🔧 2024-07-19
Ling-3.0-flash Inclusionai 262K $0.02 $0.06 🧠 🔧 2026-07-23
Nex-N2-Mini Nex AGI 262K $0.03 $0.10 🧠 🖼️ 🔧 2026-06-24
gpt-oss-20b OpenAI 131K $0.03 $0.13 🧠 🔧 2025-08-05
DeepSeek V4 Flash Latest Deepseek 1,311K $0.03 $0.16 🧠 🔧 2026-08-01
Solar Pro 4 Upstage 524K $0.03 $0.12 🧠 🔧 2026-08-10
Qwen3.7 Flash Qwen 1,000K $0.03 $0.13 🧠 🖼️ 🔧 2026-07-27
Nova Micro 1.0 Amazon 128K $0.04 $0.14 🔧 2024-12-05
gpt-oss-120b OpenAI 131K $0.04 $0.17 🧠 🔧 2025-08-05
Command R7B (12-2024) Cohere 128K $0.04 $0.15 2024-12-14
Qwen3 30B A3B Instruct 2507 Qwen 262K $0.05 $0.19 🔧 2025-07-29
🧠 reasoning · 🖼️ multimodal · 🔧 tools · Data: OpenRouter

Source: OpenRouterAPI documentation · Data as of 2026-08-30 17:07 UTC · source code

Open LLMs

217

Permissive license

80%

Gated models

11

Multimodal

142

New this week

28

Traction against release date — bubble size is 30-day downloads
TextMultimodalVision / OCRbubble = 30-day downloads

Far right and small = published days ago and already liked, before the download counter can move. That corner is invisible to any table ranked by downloads. Third-party quantizations are excluded here and kept in the catalog below; the likes axis is logarithmic.

Just landed
Published in the last 45 days, ranked by likes — the newcomers
Model Org Type Downloads 30d Likes License Gated Published
Qwen3.8-27B Qwen Multimodal 4.5M 13318 apache-2.0 2026-08-05
Qwen3.8-Flash-Next Qwen Multimodal 122k 4357 other 2026-08-24
DeepSeek-V4-Flash-0731 deepseek-ai Text 4.6M 3817 mit 2026-07-31
Muse-Glimmer-30B meta-models Multimodal 592k 1808 apache-2.0 2026-08-09
GLM-5.3-Flash zai-org Text 347k 1684 mit 2026-08-25
GLM-5.3 zai-org Text 50k 1325 other 2026-08-25
Qwen3.8-2.4T-A95B Qwen Text 35k 1183 other 2026-08-08
Qwen3.8-27B-OBLITERATED OBLITERATUS Text 726k 940 apache-2.0 2026-08-19
DeepSeek-V4-Pro-0813 deepseek-ai Text 127k 784 mit 2026-08-13
LFM2.5-2.6B LiquidAI Text 196k 718 other 2026-07-28
Ornith-1.5-35B-A3B ornith-ai Text 147k 502 mit 2026-08-18
Hy4-preview tencent Text 2k 311 apache-2.0 2026-08-27
s1-mini superwhisper Text 6k 307 other 2026-08-12
Ornith-1.5-9B ornith-ai Text 200k 255 mit 2026-08-18
needle2 Cactus-Compute Text 41k 253 apache-2.0 2026-07-29
🔒 gated (acceptance required) · Data: Hugging Face Hub
Most-downloaded open models
Ranked by 30-day downloads — established adoption, slow to move
Model Org Type Downloads 30d Likes License Gated Published
Qwen3-0.6B Qwen Text 22.5M 1558 apache-2.0 2025-04-27
gpt2 openai-community Text 14.4M 3441 mit 2022-03-02
Qwen3-8B Qwen Text 13.7M 1330 apache-2.0 2025-04-27
Qwen3.5-9B Qwen Multimodal 12.6M 1878 apache-2.0 2026-02-27
opt-125m facebook Text 11.4M 294 other 2022-05-11
Qwen3.6-35B-A3B-NVFP4 nvidia Text 11.2M 575 apache-2.0 2026-05-27
Qwen2.5-7B-Instruct Qwen Text 10.8M 1568 apache-2.0 2024-09-16
Qwen3-VL-8B-Instruct Qwen Multimodal 8.6M 1073 apache-2.0 2025-10-11
gemma-4-31B-it google Multimodal 8.5M 3670 apache-2.0 2026-03-11
Qwen2.5-VL-7B-Instruct Qwen Multimodal 8.3M 1688 apache-2.0 2025-01-26
gemma-4-26B-A4B-it google Multimodal 8.2M 1455 apache-2.0 2026-03-11
Qwen2.5-1.5B-Instruct Qwen Text 7.9M 811 apache-2.0 2024-09-17
Qwen3.5-4B Qwen Multimodal 7.6M 863 apache-2.0 2026-02-27
Qwen2.5-3B-Instruct Qwen Text 7.5M 556 other 2024-09-17
OTel-2.0-LLM-31B-IT farbodtavakkoli Text 6.8M 14 apache-2.0 2026-07-23
🔒 gated (acceptance required) · Data: Hugging Face Hub
Community favorites
Same set, ranked by likes — surfaces the flagship models
Model Org Type Downloads 30d Likes License Gated Published
DeepSeek-R1 deepseek-ai Text 2.2M 13598 mit 2025-01-20
Qwen3.8-27B Qwen Multimodal 4.5M 13318 apache-2.0 2026-08-05
Kimi-K3 moonshotai Multimodal 2.8M 11094 other 2026-06-13
Llama-3.1-8B-Instruct meta-llama Text 5.9M 6705 llama3.1 🔒 2024-07-18
gpt-oss-120b openai Text 5.3M 5136 apache-2.0 2025-08-04
gpt-oss-20b openai Text 6.5M 4969 apache-2.0 2025-08-04
Qwen3.8-Flash-Next Qwen Multimodal 122k 4357 other 2026-08-24
Unlimited-OCR baidu Multimodal 3.1M 4150 mit 2026-06-19
DeepSeek-V4-Flash-0731 deepseek-ai Text 4.6M 3817 mit 2026-07-31
gemma-4-31B-it google Multimodal 8.5M 3670 apache-2.0 2026-03-11
gpt2 openai-community Text 14.4M 3441 mit 2022-03-02
DeepSeek-OCR deepseek-ai Multimodal 2.4M 3348 mit 2025-10-17
LocateAnything-3B nvidia Multimodal 95k 2972 other 2026-03-02
Qwen3.6-35B-A3B Qwen Multimodal 5.0M 2752 apache-2.0 2026-04-15
Qwen3.6-27B Qwen Multimodal 5.8M 2282 apache-2.0 2026-04-21
🔒 gated (acceptance required) · Data: Hugging Face Hub

Source: Hugging Face HubAPI documentation · Data as of 2026-08-30 17:07 UTC · source code

What LLM Radar tracks

LLM Radar is an open dashboard that consolidates the large-language-model landscape into one continuously-updated view, split into two catalogs:

  • Hosted — commercial models you call through an API. Sourced from the public OpenRouter /api/v1/models endpoint: pricing per million tokens (input/output), context-window size, modalities, and release date.
  • Open-weights — models you can download and self-host. Sourced from the public Hugging Face Hub model listing: 30-day downloads, community likes, license, gated status, and type.

How the open-weights universe is chosen. Two decisions shape what appears, and both exist to stop the catalog from lagging the market:

  • Not only text. Open models are no longer just text-generation; a new multimodal release lands under image-text-to-text and would be invisible to a text-only query. Each tracked type is queried separately and labelled.
  • Not only downloads. The 30-day download counter takes weeks to move, so a ranking built on it can only show what was already popular — a model published two days ago sits at zero however important it is. Each type is therefore also pulled by the Hub’s trending signal, which surfaces a release the day after it lands. Both lists are merged.

The data refreshes daily via a GitHub Actions cron that re-fetches both public APIs, re-renders the dashboard, and redeploys — no server, no backend, no API key.

Both catalogs last refreshed 2026-08-30 17:07 UTC.
Definitions
  • Context window (K) — the maximum number of tokens (prompt + response) a model can process in one call, in thousands. A 200K context ≈ ~150,000 words.
  • $/1M in · $/1M out — price in US dollars per one million input / output tokens, as published by the provider on OpenRouter.
  • Permissive license — an open-weights license (Apache-2.0, MIT, and similar) that allows commercial use and redistribution without acceptance gates. The opposite is a restricted or research-only license.
  • Gated model — an open-weights model whose download requires accepting terms or requesting access. Relevant for data-residency and commercial-use decisions.
  • Cheapest capable — lowest input price among models with ≥100K context, i.e. the cheapest model still large enough for real document / agent workloads.
  • Type — what a model can be fed. On the hosted side it is derived from the input modalities OpenRouter publishes (text only · image · document + image · audio / video); on the open-weights side it is the Hub pipeline tag (text · multimodal · vision / OCR). It is read from the API, never assigned by hand.
  • Just landed — models published in the last 45 days, ranked by likes. Ranking recent arrivals by downloads would return an empty table: the counter has not had time to move.
  • Third-party quantization — a repackaging of someone else’s weights (GGUF, AWQ, MLX, 4-bit and similar). Real downloads, but not a new model, and numerous enough to crowd out the originals. Kept in the searchable catalog, excluded from the charts and the ranked tables.
  • Open weights available (hosted chart) — the hosted model has a published Hugging Face repository, so the same weights can be self-hosted. Shown as a filled marker; closed models are hollow.
Methodology & how to cite

Pipeline: OpenRouter API + Hugging Face API → Python (pandas, great-tables) → Quarto dashboard → GitHub Pages, refreshed daily by GitHub Actions. Fully open-source and reproducible — the entire pipeline is in the source repository.

LLM Radar is an independent project and is not affiliated with OpenRouter or Hugging Face. Figures reflect what those public APIs report at fetch time; always confirm pricing and licenses against the provider before relying on them.

To cite: LLM Radar — the full LLM landscape (hosted pricing & context plus open-weights adoption & licenses), by Ronald Mego. https://ronaldmego.github.io/llm-radar