Path A — Existing PC
Use what you own.
Most people underestimate their current GPU. A 12–32 GB card runs excellent quantized models every day, free of charge.
- + 12–32 GB VRAM class
- + Ollama / LM Studio ready
- + Zero extra cost
The AI ecosystem ships faster than anyone can read. We track model momentum, local hardware fit and compute prices, then turn them into one clear answer.
“What should I actually use today?”
One decision, not 50 tabs.
Since the previous edition · newest reading 2026-09-18 11:51 UTC
Pick of the week — Week 37, 2026
Gemma 4 12B
Gemma 4 12B is this week's pick by rule: the largest counted Heat Score rise among tracked models that run comfortably on a consumer card, from 45 to 62 since the evening of 7 September. Its trending score is back where it stood in late August, its measured Q4_K_M file is 6.6 GB, and it runs with room to spare on any 16 GB card at 8K context.
00 / Radar view
A contact appears at the center when its Hugging Face repository is created and drifts outward as it ages. Brightness is its Heat Score. Bearing groups models by maker and means nothing else.
How to read the scope
Newest contacts
Strongest returns
Ages come from Hugging Face repository creation dates; the Heat Score from measured signals. How the Heat Score is measured
01 / Model market
| Rank | Model | Heat0–100 · composite | TrendingHugging Face | DownloadsHugging Face | $/M inOpenRouter |
|---|---|---|---|---|---|
| 1 | Qwen3.8 27BAlibaba Qwen · OpenGeneral · Coding · Agents · ResearchHugging Face ↗ (External link)GGUF ↗ (External link)Trending: 608$/M in: $0.21 | 608 | 7.5M+1.8% rising | $0.21 | |
| 2 | Llama 3.1 8BMeta · OpenGeneral · CodingHugging Face ↗ (External link)GGUF ↗ (External link)Trending: 216 | 216 | 5.9M+4.6% rising | — | |
| 3 | Gemma 4 31BGoogle · OpenGeneral · ResearchHugging Face ↗ (External link)GGUF ↗ (External link)Trending: 50$/M in: $0.09 | 50 | 9.1M+4.2% rising | $0.09 | |
| 4 | Qwen3.8 Flash-NextAlibaba Qwen · OpenGeneral · Coding · AgentsHugging Face ↗ (External link)GGUF ↗ (External link)Trending: 226$/M in: $0.15 | 226 | 706K+25.2% rising | $0.15 | |
| 5 | Qwen3 8BAlibaba Qwen · OpenGeneral · Coding · AgentsHugging Face ↗ (External link)GGUF ↗ (External link)Trending: 62.5$/M in: $0.12 | 62.5 | 13.1M+1.4% rising | $0.12 | |
| 6 | Llama 3.2 3BMeta · OpenGeneralHugging Face ↗ (External link)GGUF ↗ (External link)Trending: 29$/M in: $0.05 | 29 | 1.8M+17.4% rising | $0.05 | |
| 7 | GLM-5.3-FlashBreakoutZhipu AI · OpenGeneral · Coding · AgentsHugging Face ↗ (External link)GGUF ↗ (External link)Trending: 163$/M in: $0.09 | 163 | 2.4M+139.1% rising | $0.09 | |
| 8 | Gemma 4 26B-A4BGoogle · OpenGeneral · ResearchHugging Face ↗ (External link)GGUF ↗ (External link)Trending: 18$/M in: $0.09 | 18 | 9.6M+8.4% rising | $0.09 | |
| 9 | Ornith 1.5 35B-A3BOrnith AI · OpenGeneral · Coding · AgentsHugging Face ↗ (External link)GGUF ↗ (External link)Trending: 43 | 43 | 394K+25.0% rising | — | |
| 10 | GLM-5.3CoolingZhipu AI · OpenGeneral · Agents · ResearchHugging Face ↗ (External link)GGUF ↗ (External link)Trending: 47$/M in: $1.40 | 47 | 838K+51.9% rising | $1.40 | |
| 11 | gpt-oss 20BOpenAI · OpenGeneral · Agents · ResearchHugging Face ↗ (External link)GGUF ↗ (External link)Trending: 26$/M in: $0.03 | 26 | 6.7M+1.3% rising | $0.03 | |
| 12 | Ornith 1.5 9BOrnith AI · OpenGeneral · CodingHugging Face ↗ (External link)GGUF ↗ (External link)Trending: 21 | 21 | 506K+35.4% rising | — |
Top 12 of 46 tracked models, ranked by Heat Score; models still collecting history follow, ranked by trending.
Heat is a composite of measured signals only: each component is the model's percentile among models with a full week of history — trending level (35%), 7-day download growth (30%), 7-day trending change (15%), 30-day downloads (15%) and Hub likes (5%). Missing components renormalize the weights and lower the shown confidence; nothing is guessed. Trending and downloads come from the Hugging Face Hub — the download counter is a rolling 30-day window, not unique users — and input pricing from OpenRouter (CC BY 4.0). The full formula, thresholds and flag rules are on the methodology page. How Heat is computed → Every tracked model has its own page →
02 / Compute ladder
Models need machines. The right machine depends on how often you push past your own hardware — not on the biggest number on a spec sheet.
Path A — Existing PC
Most people underestimate their current GPU. A 12–32 GB card runs excellent quantized models every day, free of charge.
Path B — Personal AI box
DGX Spark and GB10-class systems put 128 GB of unified memory on your desk. Rational when large local models are your daily routine.
Path C — Elastic compute
Occasional heavy job? Rent an H100 for the afternoon instead of buying hardware that idles the rest of the year.
03 / My setup
Local first. — Fit score 73
Local first.
Why this route
Hugging Face ↗ (External link)
Estimated from quantized model file sizes plus runtime overhead and a safety margin. Real usage varies with context length.

Tap the bear to wake it
From the channel
The same numbers as on this site in videos under a minute: does this model fit that card, and why. Measured file sizes, computed context memory, no speed claims.
04 / Hardware radar
Modelsneedmachines.
We track the machines that matter for local AI — consumer GPUs, personal AI systems and the rentable heavy metal.
Hardware watchlist
| Device | Memory | Class | Signal |
|---|---|---|---|
| RTX 4070Entry local | 12 GB | Steadysteady | |
| RTX 3060 12GBBudget classic | 12 GB | Huge install basesteady | |
| RTX 4080Solid local | 16 GB | Steadysteady | |
| RTX 5080Current 16 GB | 16 GB | Current genrising | |
| RTX 4060 Ti 16GBBudget 16 GB | 16 GB | Value picksteady | |
| RTX 5060 Ti 16GBCurrent budget 16 GB | 16 GB | Current genrising | |
| RTX 5070 TiFast 16 GB | 16 GB | Current genrising | |
| RTX 3090Used-market value | 24 GB | Demand risingrising | |
| RTX 4090Local workhorse | 24 GB | Demand risingrising | |
| RTX 5090Top consumer | 32 GB | Demand risingrising | |
| Mac mini M4 Pro 64GBApple unified 64 GB | 64 GB U | Apple Siliconrising | |
| Mac Studio M4 Max 64GBApple unified 64 GB | 64 GB U | Apple Siliconrising | |
| Mac Studio M3 Ultra 96GBApple unified 96 GB | 96 GB U | Apple Siliconrising | |
| NVIDIA DGX SparkPersonal AI system | 128 GB U | New categoryrising | |
| A100 80 GB (rental)Rental workhorse | 80 GB | $/h fallingfalling | |
| H100 80 GB (rental)Rental flagship | 80 GB | $/h trackedrising | |
| B200 180 GB (rental)Rental frontier | 180 GB | $/h fallingfalling |
Radar Card of the day
In 2022 this machine shipped with 24 GB. Today 22 of 46 tracked models run on it at 8K context.