Models that run on 32 GB
986 models with published weights that fit in 32 GB — a well-specified laptop. Counted at four-bit quantisation, weights only, leaving the machine about a third of its memory and a gigabyte for context: room for roughly 33 billion parameters. A mixture of experts is counted in full, because it is held in full even though only a few experts compute. 986 in all, a hundred to a page; this is page 4 of 10.
Enough for a thirty-billion-parameter model at four-bit with a long context, or a smaller one at higher precision if quality matters more than size. A practical ceiling for a laptop that also has to be a laptop.
- VibeVoice-ASR-Streaming-7BMicrosoft8.7B≈5.6 GB at 4-bit
- Molmo2-8BAllen Institute for AI (Ai2)8.7B≈5.6 GB at 4-bit
- codegemma-7bGoogle8.5B≈5.5 GB at 4-bit1 also selling it hosted
- gemma-7bGoogle8.5B≈5.5 GB at 4-bit1 also selling it hosted
- gemma-7b-itGoogle8.5B≈5.5 GB at 4-bit3 also selling it hosted
- jetmoe-8bJetMoE8.5B≈5.5 GB at 4-bit
- Emu3-GenBAAI8.5B≈5.5 GB at 4-bit
- MOSS-TTSOpenMOSS-Team8.5B≈5.5 GB at 4-bit
- MOSS-TTS-v1.5OpenMOSS-Team8.5B≈5.5 GB at 4-bit
- LFM2.5-8B-A1BLiquid AI8.5B≈5.5 GB at 4-bit
- idefics2-8bHuggingFaceM48.4B≈5.5 GB at 4-bit
- personaplex-7b-v1NVIDIA8.4B≈5.4 GB at 4-bit
- LFM2-8B-A1BLiquid AI8.3B≈5.4 GB at 4-bit
- Cosmos-Reason1-7BNVIDIA8.3B≈5.4 GB at 4-bit
- Fara-7BMicrosoft8.3B≈5.4 GB at 4-bit
- Holo1-7BH Company8.3B≈5.4 GB at 4-bit
- NuMarkdown-8B-ThinkingNuMind8.3B≈5.4 GB at 4-bit
- Qwen2.5-VL-7B-InstructAlibaba8.3B≈5.4 GB at 4-bit1 also selling it hosted
- RolmOCRReducto8.3B≈5.4 GB at 4-bit1 also selling it hosted
- Qwen2-VL-7B-InstructAlibaba8.3B≈5.4 GB at 4-bit1 also selling it hosted
- UI-TARS-7B-DPOByteDance Seed8.3B≈5.4 GB at 4-bit
- olmOCR-7B-0225-previewAi28.3B≈5.4 GB at 4-bit
- zeta-2Zed Industries8.3B≈5.4 GB at 4-bit
- VLM_WebSight_finetunedHuggingFaceM48.2B≈5.3 GB at 4-bit
- II-Medical-8BIntelligent Internet8.2B≈5.3 GB at 4-bit
- Qwen3 8BAlibaba8.2B≈5.3 GB at 4-bit8 also selling it hosted
- MisoTTSMiso Labs8.2B≈5.3 GB at 4-bit
- MiniCPM4.1-8BOpenBMB8.2B≈5.3 GB at 4-bit
- granite-3.0-8b-instructIBM Granite8.2B≈5.3 GB at 4-bit1 also selling it hosted
- Flex.2-previewostris8.2B≈5.3 GB at 4-bit
- Flex.1-alphaostris8.2B≈5.3 GB at 4-bit
- flux.1-lite-8B-alphaFreepik8.2B≈5.3 GB at 4-bit
- Qwen3-VL-Embedding-8BAlibaba8.1B≈5.3 GB at 4-bit
- ERNIE-ImageBaidu8.0B≈5.2 GB at 4-bit
- ERNIE-Image-TurboBaidu8.0B≈5.2 GB at 4-bit
- dolphin-2.9-llama3-8bDolphin8.0B≈5.2 GB at 4-bit
- UserLM-8bMicrosoft8.0B≈5.2 GB at 4-bit
- Llama3-ChatQA-1.5-8BNVIDIA8.0B≈5.2 GB at 4-bit
- Aion-RP 1.0 (8B)AionLabs8.0B≈5.2 GB at 4-bit2 also selling it hosted
- DarkIdol-Llama-3.1-8B-Instruct-1.2-Uncensoredaifeifei7988.0B≈5.2 GB at 4-bit
- KernelLLMfacebook8.0B≈5.2 GB at 4-bit
- Llama 3.1 8B InstructMeta8.0B≈5.2 GB at 4-bit20 also selling it hosted
- Llama-3-8B-Instruct-Gradient-1048kDeepSky8.0B≈5.2 GB at 4-bit
- Llama-3-8B-WebMcGill NLP Group8.0B≈5.2 GB at 4-bit
- Llama-3-RefueledRefuel AI8.0B≈5.2 GB at 4-bit
- Llama-3.1-Nemotron-Nano-8B-v1NVIDIA8.0B≈5.2 GB at 4-bit
- Llama-3.1-SuperNova-LiteArcee AI8.0B≈5.2 GB at 4-bit
- Llama3-8B-Chinese-Chatshenzhi-wang8.0B≈5.2 GB at 4-bit
- Meta-Llama-3.1-8B-Instruct-abliteratedmlabonne8.0B≈5.2 GB at 4-bit
- llama-3-Korean-Bllossom-8BMLP-LAB8.0B≈5.2 GB at 4-bit
- aya-23-8BCohere Labs8.0B≈5.2 GB at 4-bit
- c4ai-command-r7b-12-2024Cohere Labs8.0B≈5.2 GB at 4-bit
- Molmo-7B-D-0924Ai28.0B≈5.2 GB at 4-bit
- LLaDA-8B-InstructGSAI-ML8.0B≈5.2 GB at 4-bit
- Apertus-8B-2509swiss-ai8.0B≈5.2 GB at 4-bit
- Apertus-8B-Instruct-2509swiss-ai8.0B≈5.2 GB at 4-bit1 also selling it hosted
- Apertus-v1.5-8Bswiss-ai8.0B≈5.2 GB at 4-bit1 also selling it hosted
- Bio-Medical-MultiModal-Llama-3-8B-V1ContactDoctor8.0B≈5.2 GB at 4-bit
- DeepSeek R1 0528 Qwen3 8BDeepSeek8.0B≈5.2 GB at 4-bit1 also selling it hosted
- DeepSeek-R1-Distill-Llama-8BDeepSeek8.0B≈5.2 GB at 4-bit3 also selling it hosted
- F2LLM-v2-8Bcodefuse-ai8.0B≈5.2 GB at 4-bit
- Foundation-Sec-8Bfdtn-ai8.0B≈5.2 GB at 4-bit
- Hermes 2 Pro Llama 3 8BNous Research8.0B≈5.2 GB at 4-bit2 also selling it hosted
- Idefics3-8B-Llama3HuggingFaceM48.0B≈5.2 GB at 4-bit
- L3 8B Stheno V3.2Sao10K8.0B≈5.2 GB at 4-bit4 also selling it hosted
- L3-8B-Lunaris-v1Sao10K8.0B≈5.2 GB at 4-bit2 also selling it hosted
- Llama-3-8B-Lexi-UncensoredOrenguteng8.0B≈5.2 GB at 4-bit
- Llama-3-ELYZA-JP-8Belyza8.0B≈5.2 GB at 4-bit
- Llama-3-Groq-8B-Tool-UseGroq8.0B≈5.2 GB at 4-bit
- Llama-3-Open-Ko-8Bbeomi8.0B≈5.2 GB at 4-bit
- Llama-3.1-8B-Lexi-Uncensored-V2Orenguteng8.0B≈5.2 GB at 4-bit1 also selling it hosted
- Llama-3.1-Nemotron-Nano-VL-8B-V1NVIDIA8.0B≈5.2 GB at 4-bit
- Llama-3.1-Storm-8Bakjindal532448.0B≈5.2 GB at 4-bit
- Llama-3.1-Tulu-3-8BAllen Institute for AI (Ai2)8.0B≈5.2 GB at 4-bit
- Llama-Guard-3-8BMeta8.0B≈5.2 GB at 4-bit4 also selling it hosted
- Llama3-OpenBioLLM-8Baaditya8.0B≈5.2 GB at 4-bit
- Llama3-TAIDE-LX-8B-Chat-Alpha1taide8.0B≈5.2 GB at 4-bit
- Meta-Llama-Guard-2-8BMeta Llama8.0B≈5.2 GB at 4-bit1 also selling it hosted
- MiniCPM4-8BOpenBMB8.0B≈5.2 GB at 4-bit
- Mistral-NeMo-Minitron-8B-BaseNVIDIA8.0B≈5.2 GB at 4-bit
- NeuralDaredevil-8B-abliteratedmlabonne8.0B≈5.2 GB at 4-bit
- Octen-Embedding-8BOcten8.0B≈5.2 GB at 4-bit
- Qwen3 8B (embeddings)Alibaba8.0B≈5.2 GB at 4-bit6 also selling it hosted
- Qwen3-Reranker-8BAlibaba8.0B≈5.2 GB at 4-bit2 also selling it hosted
- Seed-Coder-8B-ReasoningByteDance Seed8.0B≈5.2 GB at 4-bit
- SenseNova-U1-8B-MoTsensenova8.0B≈5.2 GB at 4-bit
- WeDLM-8B-InstructTencent Hunyuan8.0B≈5.2 GB at 4-bit
- aya-vision-8bCohere Labs8.0B≈5.2 GB at 4-bit
- deepthought-8b-llama-v0.01-alpharuliad8.0B≈5.2 GB at 4-bit
- granite-3.1-8b-instructIBM Granite8.0B≈5.2 GB at 4-bit
- granite-3.3-8b-instructIBM Granite8.0B≈5.2 GB at 4-bit1 also selling it hosted
- granite-speech-3.3-8bIBM Granite8.0B≈5.2 GB at 4-bit
- llama-3-sqlcoder-8bdefog8.0B≈5.2 GB at 4-bit
- llama-embed-nemotron-8bNVIDIA8.0B≈5.2 GB at 4-bit
- Gemma-4-E4BGoogle8.0B≈5.2 GB at 4-bit
- Ling-3.0-tinyInclusionAI7.9B≈5.1 GB at 4-bit
- NV-Embed-v2NVIDIA7.9B≈5.1 GB at 4-bit
- EXAONE-3.0-7.8B-InstructLG AI Research7.8B≈5.1 GB at 4-bit
- phixtral-4x2_8mlabonne7.8B≈5.1 GB at 4-bit
- EXAONE-3.5-7.8B-InstructLGAI-EXAONE7.8B≈5.1 GB at 4-bit