LIVE
Mistral AI Raises €3 Billion to Push Open-Weight Models to the Frontier08/09/26 · Mistral AI|OpenAI agents hijacked a German wiki before the Hugging Face breach07/09/26 · OpenAI|OpenAI's GPT-6 Astra sets new state-of-the-art on ARC-AGI-303/09/26 · OpenAI|OpenAI Launches $1B Daybreak Program for Frontline Cybersecurity Defenders03/09/26 · OpenAI|Meta launches Muse Spark 1.3, a model built for agentic coding02/09/26 · Meta|Post-trained Nemotron model outscores the top human at the International Olympiad in Informatics02/09/26 · NVIDIA|Google DeepMind launches Fairwind, a proactive cyber defense program for governments and enterprises02/09/26 · Google DeepMind|Study suggests neural networks implicitly encode symbolic structures02/09/26|LibreOffice sees download surge after pledging to stay AI-free08/09/26 · LibreOffice|Danijar Hafner is building agents designed to plan for the unexpected08/09/26|OpenAI launches $5 million grant program on AI and teen development08/09/26 · OpenAI|TradingAgents: an open-source framework simulates a trading desk with LLM agents08/09/26|Mistral AI Raises €3 Billion to Push Open-Weight Models to the Frontier08/09/26 · Mistral AI|OpenAI agents hijacked a German wiki before the Hugging Face breach07/09/26 · OpenAI|OpenAI's GPT-6 Astra sets new state-of-the-art on ARC-AGI-303/09/26 · OpenAI|OpenAI Launches $1B Daybreak Program for Frontline Cybersecurity Defenders03/09/26 · OpenAI|Meta launches Muse Spark 1.3, a model built for agentic coding02/09/26 · Meta|Post-trained Nemotron model outscores the top human at the International Olympiad in Informatics02/09/26 · NVIDIA|Google DeepMind launches Fairwind, a proactive cyber defense program for governments and enterprises02/09/26 · Google DeepMind|Study suggests neural networks implicitly encode symbolic structures02/09/26|LibreOffice sees download surge after pledging to stay AI-free08/09/26 · LibreOffice|Danijar Hafner is building agents designed to plan for the unexpected08/09/26|OpenAI launches $5 million grant program on AI and teen development08/09/26 · OpenAI|TradingAgents: an open-source framework simulates a trading desk with LLM agents08/09/26|
Hardware Match

Which model for RTX 4060 · 8 Go?

8 GB usable for the model. 49 open-weight models fit in memory.

The margin goes to context: count ~1 GB per 8k tokens for a 7B model. A small margin means a short context.

ModelEloQuantizationRAM requiredMarginSpeedRun
Gemma 3n E4BGoogle · 8B1317Q4_K_M7 Go+1 Gotight ollama run gemma3n:e4b
Granite 4.1 8BIBM · 8B1306Q4_K_M7 Go+1 GotightHugging Face ↗
Gemma 3 4BGoogle · 4B1303Q8_07 Go+1 Gotight ollama run gemma3:4b
Ministral 8BMistral AI · 8.02B1237Q4_K_M7 Go+1 GotightHugging Face ↗
Llama 3.1 8BMeta · 8.03B1211Q4_K_M7 Go+1 Gotight ollama run llama3.1:8b
Llama 3.2 3BMeta · 3.21B1166Q8_06 Go+2 Gotight ollama run llama3.2:3b
Llama 3.2 1BMeta · 1.24B1111FP165 Go+3 Gocomfortable ollama run llama3.2:1b
Mistral 7BMistral AI · 7.25B1110Q4_K_M7 Go+1 Gotight ollama run mistral:7b
Kokoro-82MHexgrad · 0.08BFP162 Go+6 GocomfortableHugging Face ↗
ChatterboxResemble AI · 0.5BFP163 Go+5 GocomfortableHugging Face ↗
Parakeet TDT 0.6B v3NVIDIA · 0.6BFP163 Go+5 GocomfortableHugging Face ↗
Qwen3 0.6BAlibaba · 0.6BFP163 Go+5 Gocomfortable ollama run qwen3:0.6b
TripoSRStability AI · 0.5BFP163 Go+5 GocomfortableHugging Face ↗
Gemma 3 1BGoogle · 1BFP164 Go+4 Gocomfortable ollama run gemma3:1b
Stable Fast 3DStability AI · 1BFP164 Go+4 GocomfortableHugging Face ↗
TRELLISMicrosoft · 1.2BFP165 Go+3 GocomfortableHugging Face ↗
Whisper large-v3OpenAI · 1.55BFP165 Go+3 GocomfortableHugging Face ↗
Dia 1.6BNari Labs · 1.6BFP166 Go+2 GotightHugging Face ↗
Gemma 4 E2BGoogle · 2BFP166 Go+2 GotightHugging Face ↗
Granite 4.1 3BIBM · 3BQ8_06 Go+2 GotightHugging Face ↗
Hunyuan3D 2.1Tencent · 3BQ8_06 Go+2 GotightHugging Face ↗
Lucie 7BOpenLLM France · 6.7BQ4_K_M6 Go+2 GotightHugging Face ↗
Ministral 3 3BMistral AI · 3BQ8_06 Go+2 GotightHugging Face ↗
OLMo 3 7BAllen AI · 7BQ4_K_M6 Go+2 GotightHugging Face ↗
Orpheus 3BCanopy Labs · 3BQ8_06 Go+2 GotightHugging Face ↗
SmolLM3 3BHugging Face · 3.1BQ8_06 Go+2 GotightHugging Face ↗
Voxtral Mini 3BMistral AI · 3BQ8_06 Go+2 GotightHugging Face ↗
Z-Image TurboAlibaba · 6BQ4_K_M6 Go+2 GotightHugging Face ↗
Apertus 8BSwiss AI · 8BQ4_K_M7 Go+1 GotightHugging Face ↗
CodeGemma 7BGoogle · 8.5BQ4_K_M7 Go+1 Gotight ollama run codegemma
Command R7BCohere · 8BQ4_K_M7 Go+1 Gotight ollama run command-r7b
DeepSeek-R1 Distill 8BDeepSeek · 8.03BQ4_K_M7 Go+1 Gotight ollama run deepseek-r1:8b
FLUX.2 [klein] 4BBlack Forest Labs · 4BQ8_07 Go+1 GotightHugging Face ↗
Gemma 4 E4BGoogle · 4BQ8_07 Go+1 GotightHugging Face ↗
Granite 3.3 8BIBM · 8.1BQ4_K_M7 Go+1 Gotight ollama run granite3.3:8b
Hermes 3 8BNous Research · 8.03BQ4_K_M7 Go+1 Gotight ollama run hermes3:8b
HunyuanVideo 1.5Tencent · 8.3BQ4_K_M7 Go+1 GotightHugging Face ↗
InternLM3 8BShanghai AI Lab · 8.8BQ4_K_M7 Go+1 GotightHugging Face ↗
Ministral 3 8BMistral AI · 8BQ4_K_M7 Go+1 GotightHugging Face ↗
Phi-4 MiniMicrosoft · 3.8BQ8_07 Go+1 Gotight ollama run phi4-mini
Qwen3 4BAlibaba · 4.02BQ8_07 Go+1 Gotight ollama run qwen3:4b
Qwen3 8BAlibaba · 8.2BQ4_K_M7 Go+1 Gotight ollama run qwen3:8b
TRELLIS.2Microsoft · 4BQ8_07 Go+1 GotightHugging Face ↗
CogVideoX 5BZhipu AI · 5BQ8_08 Go+0 GotightHugging Face ↗
EuroLLM 9BUTTER Project · 9.15BQ4_K_M8 Go+0 GotightHugging Face ↗
Falcon 3 10BTII · 10.3BQ4_K_M8 Go+0 Gotight ollama run falcon3:10b
GLM-ImageZhipu AI · 9BQ4_K_M8 Go+0 GotightHugging Face ↗
Mochi 1Genmo · 10BQ4_K_M8 Go+0 GotightHugging Face ↗
Wan 2.2 5BAlibaba · 5BQ8_08 Go+0 GotightHugging Face ↗