Models that run on 24 GB
890 models with published weights that fit in 24 GB — a 24 GB card, or a Mac with 24. Counted at four-bit quantisation, weights only, leaving the machine about a third of its memory and a gigabyte for context: room for roughly 24 billion parameters. A mixture of experts is counted in full, because it is held in full even though only a few experts compute. 890 in all, a hundred to a page; this is page 1 of 9.
A 24 GB graphics card or a Mac configured with 24. This is the first size where the well-regarded mid-weight models — the twenty-something billion parameter class — run comfortably, and where a local model starts to be a real alternative to an API for daily work rather than a demonstration.
- Voxtral Small 24B 2507Mistral AI24.3B≈15.8 GB at 4-bit4 also selling it hosted
- Dolphin Mistral 24B Venice EditionCognitive Computations24.0B≈15.6 GB at 4-bit1 also selling it hosted
- Mistral Small 3.1 24B Instruct 2503Mistral AI24.0B≈15.6 GB at 4-bit3 also selling it hosted
- Mistral-Small-3.2-24B-Instruct-2506Mistral AI24.0B≈15.6 GB at 4-bit1 also selling it hosted
- Mistral Small 3Mistral AI23.6B≈15.3 GB at 4-bit4 also selling it hosted
- sarvam-mSarvam AI23.6B≈15.3 GB at 4-bit
- solar-pro-preview-instructUpstage22.1B≈14.4 GB at 4-bit
- gpt-oss-safeguard-20bOpenAI21.5B≈14.0 GB at 4-bit6 also selling it hosted
- ERNIE-4.5-21B-A3B-PTBaidu21.0B≈13.7 GB at 4-bit1 also selling it hosted
- ERNIE-4.5-21B-A3B-ThinkingBaidu21.0B≈13.7 GB at 4-bit2 also selling it hosted
- GPT OSS 20BOpenAI20.9B≈13.6 GB at 4-bit20 also selling it hosted
- context-1chroma20.9B≈13.6 GB at 4-bit
- gpt-neox-20bEleutherAI20.7B≈13.5 GB at 4-bit
- FireRed-Image-Edit-1.0FireRed Team20.4B≈13.3 GB at 4-bit
- Qwen ImageAlibaba20.4B≈13.3 GB at 4-bit1 also selling it hosted
- Qwen-Image-2512Alibaba20.4B≈13.3 GB at 4-bit
- Qwen-Image-Edit-2509Alibaba20.4B≈13.3 GB at 4-bit
- Qwen-Image-Edit-2511Alibaba20.4B≈13.3 GB at 4-bit
- maple-previewdeepgrove20.2B≈13.1 GB at 4-bit
- Gemma-4-31B-JANG_4M-CRACKdealignai20.2B≈13.1 GB at 4-bit
- cogvlm2-llama3-chat-19Bzai-org19.5B≈12.7 GB at 4-bit
- LTX-2LTX.io18.9B≈12.3 GB at 4-bit
- cogvlm-chat-hfzai-org17.6B≈11.5 GB at 4-bit
- SenseNova-U1.5-8B-MoTsensenova17.5B≈11.4 GB at 4-bit
- Wan2.1-VACE-14BWan-AI17.3B≈11.3 GB at 4-bit
- HiDream-E1-1HiDream.ai17.1B≈11.1 GB at 4-bit
- HiDream-E1-FullHiDream.ai17.1B≈11.1 GB at 4-bit
- HiDream-I1-FullHiDream.ai17.1B≈11.1 GB at 4-bit
- Kimi-VL-A3B-ThinkingMoonshot AI16.4B≈10.7 GB at 4-bit
- Kimi-VL-A3B-Thinking-2506Moonshot AI16.4B≈10.7 GB at 4-bit
- Wan2.1-I2V-14B-480PWan-AI16.4B≈10.7 GB at 4-bit
- Wan2.1-I2V-14B-720PWan-AI16.4B≈10.7 GB at 4-bit
- LLaDA2.0-UniInclusionAI16.3B≈10.6 GB at 4-bit
- Wan2.2-S2V-14BWan-AI16.3B≈10.6 GB at 4-bit
- Ling-mini-2.0InclusionAI16.3B≈10.6 GB at 4-bit
- Ring-mini-2.0InclusionAI16.3B≈10.6 GB at 4-bit
- Instella-MoE-16B-A3B-Thinkamd16.0B≈10.4 GB at 4-bit
- deepseek-moe-16b-basedeepseek-ai16.0B≈10.4 GB at 4-bit
- deepseek-moe-16b-chatdeepseek-ai16.0B≈10.4 GB at 4-bit
- Moonlight-16B-A3B-InstructMoonshot AI16.0B≈10.4 GB at 4-bit
- starcoder2-15bbigcode16.0B≈10.4 GB at 4-bit1 also selling it hosted
- starcoderbigcode15.8B≈10.3 GB at 4-bit
- DeepSeek-Coder-V2-Lite-InstructDeepSeek15.7B≈10.2 GB at 4-bit1 also selling it hosted
- DeepSeek-V2-Litedeepseek-ai15.7B≈10.2 GB at 4-bit
- starchat-alphaHugging Face H415.5B≈10.1 GB at 4-bit
- Apriel-1.6-15b-ThinkerServiceNow-AI15.0B≈9.8 GB at 4-bit
- Phi-4-reasoning-vision-15BMicrosoft15.0B≈9.8 GB at 4-bit
- WizardCoder-15B-V1.0WizardLM Team15.0B≈9.8 GB at 4-bit
- starcoder2-15b-instruct-v0.1bigcode15.0B≈9.8 GB at 4-bit1 also selling it hosted
- Apriel-1.5-15b-ThinkerServiceNow-AI14.9B≈9.7 GB at 4-bit
- DeepCoder-14B-PreviewAgentica14.8B≈9.6 GB at 4-bit
- Qwen2.5-14B-Instruct-1MAlibaba14.8B≈9.6 GB at 4-bit
- Qwen2.5-Coder-14B-InstructAlibaba14.8B≈9.6 GB at 4-bit1 also selling it hosted
- Strand-Rust-Coder-14B-v1Fortytwo14.8B≈9.6 GB at 4-bit
- SuperNova-MediusArcee AI14.8B≈9.6 GB at 4-bit
- Qwen3 14BAlibaba14.8B≈9.6 GB at 4-bit8 also selling it hosted
- BAGEL-7B-MoTByteDance Seed14.7B≈9.5 GB at 4-bit
- Phi-4Microsoft14.7B≈9.5 GB at 4-bit6 also selling it hosted
- Qwen1.5-MoE-A2.7BAlibaba14.3B≈9.3 GB at 4-bit
- 14BCausalLM14.0B≈9.1 GB at 4-bit
- ChatTS-14Bbytedance-research14.0B≈9.1 GB at 4-bit
- DeepSeek R1 Distill QWEN 14BDeepSeek14.0B≈9.1 GB at 4-bit5 also selling it hosted
- F2LLM-v2-14Bcodefuse-ai14.0B≈9.1 GB at 4-bit
- Fathom-R1-14BFractalAIResearch14.0B≈9.1 GB at 4-bit
- Nemotron-Labs-Diffusion-14BNVIDIA14.0B≈9.1 GB at 4-bit
- Qwen2.5-14BAlibaba14.0B≈9.1 GB at 4-bit1 also selling it hosted
- Velvet-14BAlmawave14.0B≈9.1 GB at 4-bit
- WAN2.2-14B-Rapid-AllInOnePhr00t14.0B≈9.1 GB at 4-bit
- Wan-Dancer-14BWan-AI14.0B≈9.1 GB at 4-bit
- Wan2.1-T2V-14BWan-AI14.0B≈9.1 GB at 4-bit1 also selling it hosted
- rwkv-4-pile-14bBlinkDL14.0B≈9.1 GB at 4-bit
- miniGCausalLM14.0B≈9.1 GB at 4-bit
- Phi-3-medium-128k-instructMicrosoft14.0B≈9.1 GB at 4-bit1 also selling it hosted
- translategemma-12b-itGoogle13.2B≈8.6 GB at 4-bit
- NexusRaven-V2-13BNexusflow13.0B≈8.5 GB at 4-bit
- Llama-2-13b-hfMeta Llama13.0B≈8.5 GB at 4-bit
- Baichuan2-13B-ChatBaichuan Intelligent Technology13.0B≈8.5 GB at 4-bit
- CodeLlama-13b-Instruct-hfCode Llama13.0B≈8.5 GB at 4-bit
- LLaVA-13b-delta-v0liuhaotian13.0B≈8.5 GB at 4-bit
- Llama-2-13bMeta Llama13.0B≈8.5 GB at 4-bit
- Llama-2-13b-chatMeta Llama13.0B≈8.5 GB at 4-bit1 also selling it hosted
- Llama-2-13b-chat-hfMeta13.0B≈8.5 GB at 4-bit1 also selling it hosted
- Llama2-13B-TiefighterKoboldAI13.0B≈8.5 GB at 4-bit1 also selling it hosted
- Llama2-Chinese-13b-ChatFlagAlpha13.0B≈8.5 GB at 4-bit
- MythoMax 13BGryphe13.0B≈8.5 GB at 4-bit4 also selling it hosted
- Nous Hermes Llama2 13B13.0B≈8.5 GB at 4-bit2 also selling it hosted
- OpenOrca-Platypus2-13BOpenOrca13.0B≈8.5 GB at 4-bit
- Orca-2-13bMicrosoft13.0B≈8.5 GB at 4-bit
- ReMM SLERP 13BUndi9513.0B≈8.5 GB at 4-bit1 also selling it hosted
- WhiteRabbitNeo-13B-v1WhiteRabbitNeo13.0B≈8.5 GB at 4-bit
- Wizard-Vicuna-13B-UncensoredQuixi AI13.0B≈8.5 GB at 4-bit
- Wizard-Vicuna-13B-Uncensored-HFTheBloke13.0B≈8.5 GB at 4-bit
- WizardLM-13B-UncensoredQuixi AI13.0B≈8.5 GB at 4-bit
- WizardLM-13B-V1.2WizardLM Team13.0B≈8.5 GB at 4-bit
- Ziya-LLaMA-13B-v1Fengshenbang-LM13.0B≈8.5 GB at 4-bit
- chronos-hermes-13b-v2Austism13.0B≈8.5 GB at 4-bit2 also selling it hosted
- jais-13binception4213.0B≈8.5 GB at 4-bit
- jais-13b-chatinception4213.0B≈8.5 GB at 4-bit1 also selling it hosted
- llama-13bhuggyllama13.0B≈8.5 GB at 4-bit
- mythalion-13bPygmalionAI13.0B≈8.5 GB at 4-bit