Models that run on 24 GB
890 models with published weights that fit in 24 GB — a 24 GB card, or a Mac with 24. Counted at four-bit quantisation, weights only, leaving the machine about a third of its memory and a gigabyte for context: room for roughly 24 billion parameters. A mixture of experts is counted in full, because it is held in full even though only a few experts compute. 890 in all, a hundred to a page; this is page 2 of 9.
A 24 GB graphics card or a Mac configured with 24. This is the first size where the well-regarded mid-weight models — the twenty-something billion parameter class — run comfortably, and where a local model starts to be a real alternative to an API for daily work rather than a demonstration.
- open_llama_13bOpenLM Research13.0B≈8.5 GB at 4-bit
- prometheus-13b-v1.0prometheus-eval13.0B≈8.5 GB at 4-bit
- ruGPT-3.5-13Bai-forever13.0B≈8.5 GB at 4-bit
- stable-vicuna-13b-deltaCarperAI13.0B≈8.5 GB at 4-bit
- vicuna-13b-v1.5Large Model Systems Organization13.0B≈8.5 GB at 4-bit
- vicuna-13b-v1.5-16kLarge Model Systems Organization13.0B≈8.5 GB at 4-bit
- Krea-2-RawKrea12.8B≈8.3 GB at 4-bit
- Krea-2-TurboKrea12.8B≈8.3 GB at 4-bit
- Wayfarer-12BLatitude12.2B≈8.0 GB at 4-bit
- Mistral NemoMistral AI12.2B≈8.0 GB at 4-bit10 also selling it hosted
- Gemma 3 12BGoogle12.2B≈7.9 GB at 4-bit9 also selling it hosted
- Gemma-3-R1984-12BVIDraft12.2B≈7.9 GB at 4-bit
- Mellum2-12B-A2.5B-ThinkingJetBrains12.1B≈7.9 GB at 4-bit
- Llama Guard 4 12BMeta12.0B≈7.8 GB at 4-bit6 also selling it hosted
- Huihui-gemma-4-12B-coder-fable5-composer2.5-v1-abliteratedhuihui-ai12.0B≈7.8 GB at 4-bit
- KaLM-Embedding-Gemma3-12B-2511Tencent Hunyuan12.0B≈7.8 GB at 4-bit
- MN-12B-Celeste-V1.9nothingiisreal12.0B≈7.8 GB at 4-bit
- NVIDIA-Nemotron-Nano-12B-v2NVIDIA12.0B≈7.8 GB at 4-bit2 also selling it hosted
- NemoMix-Unleashed-12BMarinaraSpaghetti12.0B≈7.8 GB at 4-bit
- Pixtral 12B 2409Mistral AI12.0B≈7.8 GB at 4-bit3 also selling it hosted
- gemma-3-12b-it-qat-q4_0-unquantizedGoogle12.0B≈7.8 GB at 4-bit
- oasst-sft-1-pythia-12bOpenAssistant12.0B≈7.8 GB at 4-bit
- oasst-sft-4-pythia-12b-epoch-3.5OpenAssistant12.0B≈7.8 GB at 4-bit
- pythia-12bEleutherAI12.0B≈7.8 GB at 4-bit1 also selling it hosted
- Gemma-4-12BGoogle12.0B≈7.8 GB at 4-bit
- Gemma-4-12B-itGoogle12.0B≈7.8 GB at 4-bit
- Huihui-gemma-4-12B-it-abliteratedhuihui-ai12.0B≈7.8 GB at 4-bit
- FLUX.1-Canny-devBlack Forest Labs11.9B≈7.7 GB at 4-bit
- FLUX.1-Depth-devBlack Forest Labs11.9B≈7.7 GB at 4-bit
- AWPortrait-FLShakker Labs11.9B≈7.7 GB at 4-bit
- FLUX.1-Krea-devBlack Forest Labs11.9B≈7.7 GB at 4-bit
- PixelWave_FLUX.1-dev_03Mikey And Friends11.9B≈7.7 GB at 4-bit
- SRPOTencent Hunyuan11.9B≈7.7 GB at 4-bit
- UltraFlux-v1Owen77711.9B≈7.7 GB at 4-bit
- FLUX.1 [schnell] (Turbo)Black Forest Labs11.9B≈7.7 GB at 4-bit3 also selling it hosted
- OpenFLUX.1ostris11.9B≈7.7 GB at 4-bit
- shuttle-3-diffusionShuttleAI11.9B≈7.7 GB at 4-bit
- falcon-11BTechnology Innovation Institute11.1B≈7.2 GB at 4-bit
- Bielik-11B-v3.0-Instructspeakleash11.0B≈7.2 GB at 4-bit1 also selling it hosted
- Llama 3.2 11B Vision InstructMeta11.0B≈7.2 GB at 4-bit7 also selling it hosted
- Llama Guard 3 11B VisionMeta11.0B≈7.2 GB at 4-bit2 also selling it hosted
- Llama-3.2V-11B-cotXkev11.0B≈7.2 GB at 4-bit
- YanoljaNEXT-EEVE-Instruct-10.8Byanolja10.8B≈7.0 GB at 4-bit
- HyperCLOVAX-SEED-Omni-8BHyperCLOVA X10.7B≈7.0 GB at 4-bit
- Fimbulvetr-11B-v2Sao10K10.7B≈7.0 GB at 4-bit
- Solar-10.7B-Instruct-v1.0Upstage10.7B≈7.0 GB at 4-bit
- Solar-10.7B-v1.0Upstage10.7B≈7.0 GB at 4-bit
- Llama-3.2-11B-VisionMeta Llama10.6B≈6.9 GB at 4-bit
- Boogu-Image-0.1-EditBoogu10.3B≈6.7 GB at 4-bit
- Step3-VL-10BStepFun10.2B≈6.6 GB at 4-bit
- Kimi-Audio-7B-InstructMoonshot AI9.8B≈6.3 GB at 4-bit
- Qwen3.5 9BAlibaba9.7B≈6.3 GB at 4-bit8 also selling it hosted
- Qwen3.8-9B-Distillempero-ai9.7B≈6.3 GB at 4-bit
- liftDatalab9.7B≈6.3 GB at 4-bit
- MiniCPM-SALAOpenBMB9.5B≈6.2 GB at 4-bit
- Fun-Audio-Chat-8BQwenAudio9.5B≈6.1 GB at 4-bit
- OmniCoder-9BTesslate9.4B≈6.1 GB at 4-bit
- Qwen3.5-9B-Claude-4.6-HighIQ-THINKING-HERETIC-UNCENSOREDDavidAU9.4B≈6.1 GB at 4-bit
- Qwythos-9B-Claude-Mythos-5-1Mempero-ai9.4B≈6.1 GB at 4-bit
- fuyu-8bAdept AI Labs9.4B≈6.1 GB at 4-bit
- MiniCPM-o-4_5OpenBMB9.4B≈6.1 GB at 4-bit
- VibeVoice-7BVibeVoice Community (Unofficial)9.3B≈6.1 GB at 4-bit
- VibeVoice-Largeaoi-ot9.3B≈6.1 GB at 4-bit
- kugelaudio-0-openKugelaudio9.3B≈6.1 GB at 4-bit
- ideogram-4-fp8Ideogram9.3B≈6.0 GB at 4-bit
- moondream3-previewmoondream9.3B≈6.0 GB at 4-bit
- bge-multilingual-gemma2BAAI9.2B≈6.0 GB at 4-bit
- gemma-2-9bGoogle9.2B≈6.0 GB at 4-bit
- EuroLLM-9B-InstructUTTER - Unified Transcription and Translation for Extended Reality9.2B≈5.9 GB at 4-bit
- FLUX.2-klein-9b-kvBlack Forest Labs9.1B≈5.9 GB at 4-bit
- AutoGLM-Phone-9B-MultilingualZ.ai9.0B≈5.9 GB at 4-bit3 also selling it hosted
- EuroLLM-9BUTTER - Unified Transcription and Translation for Extended Reality9.0B≈5.9 GB at 4-bit
- FLUX.2-klein-9b-fp8Black Forest Labs9.0B≈5.9 GB at 4-bit
- FLUX.2-klein-base-9BBlack Forest Labs9.0B≈5.9 GB at 4-bit
- FLUX.2-klein-base-9b-fp8Black Forest Labs9.0B≈5.9 GB at 4-bit
- Flux2-Klein-9B-Enhanced-Detailsdx81529.0B≈5.9 GB at 4-bit
- Flux2-Klein-9B-True-V2wikeeyang9.0B≈5.9 GB at 4-bit
- Flux2-Klein-9B-True-V3wikeeyang9.0B≈5.9 GB at 4-bit
- Nemotron Nano 9B V2NVIDIA9.0B≈5.9 GB at 4-bit5 also selling it hosted
- Ornith-1.5-9Bornith-ai9.0B≈5.9 GB at 4-bit
- Ovis1.6-Gemma2-9BATH-MaaS9.0B≈5.9 GB at 4-bit
- Ovis2.5-9BATH-MaaS9.0B≈5.9 GB at 4-bit
- Qwen3.5-9B-BaseAlibaba9.0B≈5.9 GB at 4-bit1 also selling it hosted
- Qwythos-9B-v2empero-ai9.0B≈5.9 GB at 4-bit
- Turkish-Gemma-9b-T1ytu-ce-cosmos9.0B≈5.9 GB at 4-bit
- Yi-1.5-9B-Chat01-ai9.0B≈5.9 GB at 4-bit
- codegeex4-all-9bzai-org9.0B≈5.9 GB at 4-bit
- gemma-2-9b-itGoogle9.0B≈5.9 GB at 4-bit3 also selling it hosted
- gemma-2-9b-it-SimPOprinceton-nlp9.0B≈5.9 GB at 4-bit
- Carnice-9bkai-os9.0B≈5.8 GB at 4-bit
- NeoHorse-1-9BTokenRhythm9.0B≈5.8 GB at 4-bit
- Chroma1-HDlodestones8.9B≈5.8 GB at 4-bit
- Yi-9B01-ai8.8B≈5.7 GB at 4-bit
- Yi-Coder-9B-Chat01-ai8.8B≈5.7 GB at 4-bit
- internlm3-8b-instructinternlm8.8B≈5.7 GB at 4-bit
- Granite 4.1 8BIBM watsonx.ai8.8B≈5.7 GB at 4-bit
- Cosmos-Reason2-8BNVIDIA8.8B≈5.7 GB at 4-bit
- Huihui-Qwen3-VL-8B-Instruct-abliteratedhuihui-ai8.8B≈5.7 GB at 4-bit
- MAI-UI-8BTongyi-MAI8.8B≈5.7 GB at 4-bit
- Qwen3 VL 8B InstructAlibaba8.8B≈5.7 GB at 4-bit8 also selling it hosted