Models that run on 32 GB
986 models with published weights that fit in 32 GB — a well-specified laptop. Counted at four-bit quantisation, weights only, leaving the machine about a third of its memory and a gigabyte for context: room for roughly 33 billion parameters. A mixture of experts is counted in full, because it is held in full even though only a few experts compute. 986 in all, a hundred to a page; this is page 7 of 10.
Enough for a thirty-billion-parameter model at four-bit with a long context, or a smaller one at higher precision if quality matters more than size. A practical ceiling for a laptop that also has to be a laptop.
- Ovis-U1-3BATH-MaaS3.6B≈2.4 GB at 4-bit
- YuE2-3BMultimodal Art Projection3.6B≈2.4 GB at 4-bit
- Step-Audio-TTS-3Bstepfun-ai3.5B≈2.3 GB at 4-bit
- ACE-Step-v1-3.5BACE-Step3.5B≈2.3 GB at 4-bit
- NSFW-GEN-ANIMEUnfilteredAI3.5B≈2.3 GB at 4-bit
- NSFW-gen-v2UnfilteredAI3.5B≈2.3 GB at 4-bit
- Breeze-TTS-2BreezeBlue3.5B≈2.3 GB at 4-bit
- deepseek-vl2-tinydeepseek-ai3.4B≈2.2 GB at 4-bit
- Unlimited-OCRBaidu3.3B≈2.2 GB at 4-bit
- FLUX.1-dev-ControlNet-Union-ProShakker Labs3.3B≈2.1 GB at 4-bit
- nllb-200-3.3BAI at Meta3.3B≈2.1 GB at 4-bit
- Llama 3.2 3B InstructMeta3.2B≈2.1 GB at 4-bit14 also selling it hosted
- llama-3.2-Korean-Bllossom-3BBllossom3.2B≈2.1 GB at 4-bit
- Granite 4.0 MicroIBM watsonx.ai3.2B≈2.1 GB at 4-bit2 also selling it hosted
- imp-v1-3bMILVLG3.2B≈2.1 GB at 4-bit
- LFM2.5-VL-3BLiquid AI3.1B≈2.0 GB at 4-bit
- Qwen2.5-3BAlibaba3.1B≈2.0 GB at 4-bit
- Qwen2.5-3B-InstructAlibaba3.1B≈2.0 GB at 4-bit
- VibeThinker-3BWeiboAI3.1B≈2.0 GB at 4-bit
- SmolLM3-3BHugging Face Smol Models Research3.1B≈2.0 GB at 4-bit
- dots.ocrdots studio3.0B≈2.0 GB at 4-bit
- OpenELM-3B-InstructApple3.0B≈2.0 GB at 4-bit
- starcoder2-3bbigcode3.0B≈2.0 GB at 4-bit1 also selling it hosted
- Qwen2.5-Coder-3B-InstructAlibaba3.0B≈2.0 GB at 4-bit3 also selling it hosted
- SmolLM3-3B-BaseHuggingFaceTB3.0B≈2.0 GB at 4-bit
- Voxtral-Mini-3B-2507Mistral AI3.0B≈2.0 GB at 4-bit2 also selling it hosted
- bitnet_b1_58-3B1bitLLM3.0B≈2.0 GB at 4-bit
- fish-agent-v0.1-3bFish Audio3.0B≈2.0 GB at 4-bit
- open_llama_3bOpenLM Research3.0B≈2.0 GB at 4-bit
- open_llama_3b_v2OpenLM Research3.0B≈2.0 GB at 4-bit
- orca_mini_3bpankajmathur3.0B≈2.0 GB at 4-bit
- orpheus-3b-0.1-pretrainedCanopy Labs3.0B≈2.0 GB at 4-bit
- paligemma2-3b-pt-224Google3.0B≈2.0 GB at 4-bit
- proxy-lite-3bconvergence-ai3.0B≈2.0 GB at 4-bit
- replit-code-v1-3bReplit3.0B≈2.0 GB at 4-bit
- replit-code-v1_5-3bReplit3.0B≈2.0 GB at 4-bit
- stablecode-completion-alpha-3b-4kStability AI3.0B≈2.0 GB at 4-bit
- stablecode-instruct-alpha-3bStability AI3.0B≈2.0 GB at 4-bit
- stablelm-3b-4e1tStability AI3.0B≈2.0 GB at 4-bit
- tada-3b-mlHume AI3.0B≈2.0 GB at 4-bit
- madlad400-3b-mtGoogle2.9B≈1.9 GB at 4-bit
- paligemma-3b-pt-224Google2.9B≈1.9 GB at 4-bit
- Anima-2.9BGazingstars1232.9B≈1.9 GB at 4-bit
- stable-code-3bStability AI2.8B≈1.8 GB at 4-bit1 also selling it hosted
- stable-code-instruct-3bStability AI2.8B≈1.8 GB at 4-bit
- stablelm-zephyr-3bStability AI2.8B≈1.8 GB at 4-bit
- dolphin-2_6-phi-2Dolphin2.8B≈1.8 GB at 4-bit
- phi-2Microsoft2.8B≈1.8 GB at 4-bit
- Solidity-LLMChainGPT2.8B≈1.8 GB at 4-bit
- gpt-neo-2.7BEleutherAI2.7B≈1.8 GB at 4-bit
- VibeVoice-1.5BMicrosoft2.7B≈1.8 GB at 4-bit
- gemma-2-2bGoogle2.6B≈1.7 GB at 4-bit
- gemma-2-2b-itGoogle2.6B≈1.7 GB at 4-bit
- gemma-2-2b-jpn-itGoogle2.6B≈1.7 GB at 4-bit
- Lumina-Image-2.0Alpha-VLLM2.6B≈1.7 GB at 4-bit
- LFM2-2.6B-TranscriptLiquid AI2.6B≈1.7 GB at 4-bit
- Ouro-2.6B-ThinkingByteDance2.6B≈1.7 GB at 4-bit
- KolorsKolors Team, Kuaishou Technology2.6B≈1.7 GB at 4-bit
- Ovis2.5-2BATH-MaaS2.6B≈1.7 GB at 4-bit
- LFM2-2.6BLiquid AI2.6B≈1.7 GB at 4-bit
- LFM2-2.6B-ExpLiquid AI2.6B≈1.7 GB at 4-bit
- stable-diffusion-xl-1.0-inpainting-0.1🧨Diffusers2.6B≈1.7 GB at 4-bit
- Illustrious-xl-early-release-v0OnomaAI2.6B≈1.7 GB at 4-bit
- OpenDalleV1.1dataautogpt32.6B≈1.7 GB at 4-bit
- SD XLStability AI2.6B≈1.7 GB at 4-bit2 also selling it hosted
- animagine-xl-2.0Linaqruf2.6B≈1.7 GB at 4-bit
- animagine-xl-3.0Cagliostro Labs2.6B≈1.7 GB at 4-bit
- animagine-xl-3.1Cagliostro Labs2.6B≈1.7 GB at 4-bit
- animagine-xl-4.0Cagliostro Labs2.6B≈1.7 GB at 4-bit
- dpo-sdxl-text2image-v1mhdang2.6B≈1.7 GB at 4-bit
- playground-v2-1024px-aestheticPlayground2.6B≈1.7 GB at 4-bit1 also selling it hosted
- playground-v2.5-1024px-aestheticPlayground2.6B≈1.7 GB at 4-bit1 also selling it hosted
- sdxl-flashStable Diffusion Community (Unofficial, Non-profit)2.6B≈1.7 GB at 4-bit
- stable-diffusion-xl-base-0.9Stability AI2.6B≈1.7 GB at 4-bit
- MiniCPM5-2BOpenBMB2.5B≈1.6 GB at 4-bit
- Octopus-v2Nexa AI2.5B≈1.6 GB at 4-bit
- gemma-2bGoogle2.5B≈1.6 GB at 4-bit
- gemma-2b-itGoogle2.5B≈1.6 GB at 4-bit1 also selling it hosted
- BidirLM-Omni-2.5B-EmbeddingBidirLM2.5B≈1.6 GB at 4-bit
- canary-qwen-2.5bNVIDIA2.5B≈1.6 GB at 4-bit
- Cosmos-Reason2-2BNVIDIA2.4B≈1.6 GB at 4-bit
- EXAONE-3.5-2.4B-InstructLGAI-EXAONE2.4B≈1.6 GB at 4-bit
- seamless-m4t-v2-largeAI at Meta2.3B≈1.5 GB at 4-bit
- VoxCPM2OpenBMB2.3B≈1.5 GB at 4-bit
- stable-diffusion-xl-refiner-0.9Stability AI2.3B≈1.5 GB at 4-bit
- stable-diffusion-xl-refiner-1.0Stability AI2.3B≈1.5 GB at 4-bit
- GLM-ASR-Nano-2512Z.ai2.3B≈1.5 GB at 4-bit
- SmolVLM2-2.2B-InstructHugging Face Smol Models Research2.2B≈1.5 GB at 4-bit
- Marlin-2BNemo Station2.2B≈1.4 GB at 4-bit
- Qwen2-VL-2B-InstructAlibaba2.2B≈1.4 GB at 4-bit1 also selling it hosted
- tada-1bHume AI2.2B≈1.4 GB at 4-bit
- FLUX.1-dev-Controlnet-Inpainting-Betaalimama-creative2.1B≈1.4 GB at 4-bit
- FLUX.1-dev-ControlNet-Union-Pro-2.0Shakker Labs2.1B≈1.4 GB at 4-bit
- Qwen3-VL-2B-InstructAlibaba2.1B≈1.4 GB at 4-bit
- Qwen3-VL-Embedding-2BAlibaba2.1B≈1.4 GB at 4-bit
- Janus-1.3BDeepSeek2.1B≈1.4 GB at 4-bit
- stable-diffusion-3-medium-diffusersStability AI2.1B≈1.4 GB at 4-bit
- cohere-transcribe-arabic-07-2026Cohere Labs2.1B≈1.3 GB at 4-bit
- SoulX-Podcast-1.7BSoul-AILab2.1B≈1.3 GB at 4-bit
- Qwen3-1.7BAlibaba2.0B≈1.3 GB at 4-bit1 also selling it hosted