Models that run on 36 GB
1031 models with published weights that fit in 36 GB — a MacBook Pro with 36. Counted at four-bit quantisation, weights only, leaving the machine about a third of its memory and a gigabyte for context: room for roughly 37 billion parameters. A mixture of experts is counted in full, because it is held in full even though only a few experts compute. 1,031 in all, a hundred to a page; this is page 7 of 11.
An Apple configuration, and a comfortable one: it holds what 32 GB holds without the machine feeling tight, which in practice means a longer context or a browser you do not have to close first.
- Jan-nano-128kMenlo Research4.0B≈2.6 GB at 4-bit
- Jan-v1-4BJan4.0B≈2.6 GB at 4-bit
- fable-tracesAliesTaha4.0B≈2.6 GB at 4-bit
- Mellum-4b-baseJetBrains4.0B≈2.6 GB at 4-bit
- Llasa-3BHKUST Audio4.0B≈2.6 GB at 4-bit
- F2LLM-4Bcodefuse-ai4.0B≈2.6 GB at 4-bit
- F2LLM-v2-4Bcodefuse-ai4.0B≈2.6 GB at 4-bit
- FLUX.2-klein-base-4BBlack Forest Labs4.0B≈2.6 GB at 4-bit
- GELab-Zero-4B-previewstepfun-ai4.0B≈2.6 GB at 4-bit
- Gemma-3-Gaia-PT-BR-4b-itCEIA-UFG4.0B≈2.6 GB at 4-bit
- LocoOperator-4BLocoreMind4.0B≈2.6 GB at 4-bit
- LocoTrainer-4BLocoreMind4.0B≈2.6 GB at 4-bit
- MiniCPM3-4BOpenBMB4.0B≈2.6 GB at 4-bit
- Nemotron-Mini-4B-InstructNVIDIA4.0B≈2.6 GB at 4-bit
- OmniNeural-4BNexa AI4.0B≈2.6 GB at 4-bit
- Qwen3 4BAlibaba4.0B≈2.6 GB at 4-bit3 also selling it hosted
- Qwen3-4B-Instruct-2507Alibaba4.0B≈2.6 GB at 4-bit2 also selling it hosted
- Qwen3-4B-Thinking-2507Alibaba4.0B≈2.6 GB at 4-bit1 also selling it hosted
- Qwen3-4b-Z-Image-Engineer-V4BennyDaBall4.0B≈2.6 GB at 4-bit
- Qwen3-Embedding-4BAlibaba4.0B≈2.6 GB at 4-bit4 also selling it hosted
- Qwen3-Reranker-4BAlibaba4.0B≈2.6 GB at 4-bit1 also selling it hosted
- Qwen3-VL-4B-InstructAlibaba4.0B≈2.6 GB at 4-bit1 also selling it hosted
- Qwen3-VL-4B-ThinkingAlibaba4.0B≈2.6 GB at 4-bit1 also selling it hosted
- Qwen3.5-4BAlibaba4.0B≈2.6 GB at 4-bit2 also selling it hosted
- R-4BYannQi4.0B≈2.6 GB at 4-bit
- Voxtral-4B-TTS-2603Mistral AI4.0B≈2.6 GB at 4-bit
- Youtu-VL-4B-InstructTencent Hunyuan4.0B≈2.6 GB at 4-bit
- gemma-3-4b-ptGoogle4.0B≈2.6 GB at 4-bit
- medgemma-4b-ptGoogle4.0B≈2.6 GB at 4-bit
- shieldgemma-2-4b-itGoogle4.0B≈2.6 GB at 4-bit
- t5gemma-2-4b-4bGoogle4.0B≈2.6 GB at 4-bit
- OmniGen2OmniGen4.0B≈2.6 GB at 4-bit
- HeartMuLa-oss-3BHeartMuLa3.9B≈2.6 GB at 4-bit
- Nanbeige4.1-3BNanbeige LLM Lab3.9B≈2.6 GB at 4-bit
- Phi-4-mini-instructMicrosoft3.8B≈2.5 GB at 4-bit1 also selling it hosted
- LocateAnything-3BNVIDIA3.8B≈2.5 GB at 4-bit
- NuExtractNuMind3.8B≈2.5 GB at 4-bit
- NuExtract-1.5NuMind3.8B≈2.5 GB at 4-bit
- Phi-3-mini-128k-instructMicrosoft3.8B≈2.5 GB at 4-bit2 also selling it hosted
- Phi-3-mini-4k-instructMicrosoft3.8B≈2.5 GB at 4-bit1 also selling it hosted
- Phi-3.5-mini-instructMicrosoft3.8B≈2.5 GB at 4-bit1 also selling it hosted
- Hy-Embodied-0.5Tencent Hunyuan3.8B≈2.5 GB at 4-bit
- Orpheus 3BCanopy Labs3.8B≈2.5 GB at 4-bit1 also selling it hosted
- blip2-opt-2.7bSalesforce3.7B≈2.4 GB at 4-bit
- HyperCLOVAX-SEED-Vision-Instruct-3BHyperCLOVA X3.7B≈2.4 GB at 4-bit
- Ovis-U1-3BATH-MaaS3.6B≈2.4 GB at 4-bit
- YuE2-3BMultimodal Art Projection3.6B≈2.4 GB at 4-bit
- Step-Audio-TTS-3Bstepfun-ai3.5B≈2.3 GB at 4-bit
- ACE-Step-v1-3.5BACE-Step3.5B≈2.3 GB at 4-bit
- NSFW-GEN-ANIMEUnfilteredAI3.5B≈2.3 GB at 4-bit
- NSFW-gen-v2UnfilteredAI3.5B≈2.3 GB at 4-bit
- Breeze-TTS-2BreezeBlue3.5B≈2.3 GB at 4-bit
- deepseek-vl2-tinydeepseek-ai3.4B≈2.2 GB at 4-bit
- Unlimited-OCRBaidu3.3B≈2.2 GB at 4-bit
- FLUX.1-dev-ControlNet-Union-ProShakker Labs3.3B≈2.1 GB at 4-bit
- nllb-200-3.3BAI at Meta3.3B≈2.1 GB at 4-bit
- Llama 3.2 3B InstructMeta3.2B≈2.1 GB at 4-bit14 also selling it hosted
- llama-3.2-Korean-Bllossom-3BBllossom3.2B≈2.1 GB at 4-bit
- Granite 4.0 MicroIBM watsonx.ai3.2B≈2.1 GB at 4-bit2 also selling it hosted
- imp-v1-3bMILVLG3.2B≈2.1 GB at 4-bit
- LFM2.5-VL-3BLiquid AI3.1B≈2.0 GB at 4-bit
- Qwen2.5-3BAlibaba3.1B≈2.0 GB at 4-bit
- Qwen2.5-3B-InstructAlibaba3.1B≈2.0 GB at 4-bit
- VibeThinker-3BWeiboAI3.1B≈2.0 GB at 4-bit
- SmolLM3-3BHugging Face Smol Models Research3.1B≈2.0 GB at 4-bit
- dots.ocrdots studio3.0B≈2.0 GB at 4-bit
- OpenELM-3B-InstructApple3.0B≈2.0 GB at 4-bit
- starcoder2-3bbigcode3.0B≈2.0 GB at 4-bit1 also selling it hosted
- Qwen2.5-Coder-3B-InstructAlibaba3.0B≈2.0 GB at 4-bit3 also selling it hosted
- SmolLM3-3B-BaseHuggingFaceTB3.0B≈2.0 GB at 4-bit
- Voxtral-Mini-3B-2507Mistral AI3.0B≈2.0 GB at 4-bit2 also selling it hosted
- bitnet_b1_58-3B1bitLLM3.0B≈2.0 GB at 4-bit
- fish-agent-v0.1-3bFish Audio3.0B≈2.0 GB at 4-bit
- open_llama_3bOpenLM Research3.0B≈2.0 GB at 4-bit
- open_llama_3b_v2OpenLM Research3.0B≈2.0 GB at 4-bit
- orca_mini_3bpankajmathur3.0B≈2.0 GB at 4-bit
- orpheus-3b-0.1-pretrainedCanopy Labs3.0B≈2.0 GB at 4-bit
- paligemma2-3b-pt-224Google3.0B≈2.0 GB at 4-bit
- proxy-lite-3bconvergence-ai3.0B≈2.0 GB at 4-bit
- replit-code-v1-3bReplit3.0B≈2.0 GB at 4-bit
- replit-code-v1_5-3bReplit3.0B≈2.0 GB at 4-bit
- stablecode-completion-alpha-3b-4kStability AI3.0B≈2.0 GB at 4-bit
- stablecode-instruct-alpha-3bStability AI3.0B≈2.0 GB at 4-bit
- stablelm-3b-4e1tStability AI3.0B≈2.0 GB at 4-bit
- tada-3b-mlHume AI3.0B≈2.0 GB at 4-bit
- madlad400-3b-mtGoogle2.9B≈1.9 GB at 4-bit
- paligemma-3b-pt-224Google2.9B≈1.9 GB at 4-bit
- Anima-2.9BGazingstars1232.9B≈1.9 GB at 4-bit
- stable-code-3bStability AI2.8B≈1.8 GB at 4-bit1 also selling it hosted
- stable-code-instruct-3bStability AI2.8B≈1.8 GB at 4-bit
- stablelm-zephyr-3bStability AI2.8B≈1.8 GB at 4-bit
- dolphin-2_6-phi-2Dolphin2.8B≈1.8 GB at 4-bit
- phi-2Microsoft2.8B≈1.8 GB at 4-bit
- Solidity-LLMChainGPT2.8B≈1.8 GB at 4-bit
- gpt-neo-2.7BEleutherAI2.7B≈1.8 GB at 4-bit
- VibeVoice-1.5BMicrosoft2.7B≈1.8 GB at 4-bit
- gemma-2-2bGoogle2.6B≈1.7 GB at 4-bit
- gemma-2-2b-itGoogle2.6B≈1.7 GB at 4-bit
- gemma-2-2b-jpn-itGoogle2.6B≈1.7 GB at 4-bit
- Lumina-Image-2.0Alpha-VLLM2.6B≈1.7 GB at 4-bit