Models that run on 32 GB
986 models with published weights that fit in 32 GB — a well-specified laptop. Counted at four-bit quantisation, weights only, leaving the machine about a third of its memory and a gigabyte for context: room for roughly 33 billion parameters. A mixture of experts is counted in full, because it is held in full even though only a few experts compute. 986 in all, a hundred to a page; this is page 8 of 10.
Enough for a thirty-billion-parameter model at four-bit with a long context, or a smaller one at higher precision if quality matters more than size. A practical ceiling for a laptop that also has to be a laptop.
- Isaac 0.2 2B PreviewPerceptron2.0B≈1.3 GB at 4-bit1 also selling it hosted
- MOSS-Transcribe-preview-2BOpenMOSS-Team2.0B≈1.3 GB at 4-bit
- MiniCPM-2B-sft-fp32OpenBMB2.0B≈1.3 GB at 4-bit
- Qwen3.5-2BAlibaba2.0B≈1.3 GB at 4-bit2 also selling it hosted
- gemma-1.1-2b-itGoogle2.0B≈1.3 GB at 4-bit
- granite-speech-4.1-2bIBM watsonx.ai2.0B≈1.3 GB at 4-bit
- granite-speech-4.1-2b-narIBM watsonx.ai2.0B≈1.3 GB at 4-bit
- helium-1-preview-2bKyutai2.0B≈1.3 GB at 4-bit
- Youtu-LLM-2BTencent Hunyuan2.0B≈1.3 GB at 4-bit
- moondream2vikhyatk1.9B≈1.3 GB at 4-bit
- LTX-VideoLTX.io1.9B≈1.3 GB at 4-bit
- Dia2-2BNari Labs1.9B≈1.2 GB at 4-bit
- Qwen3-TTS-12Hz-1.7B-CustomVoiceAlibaba1.9B≈1.2 GB at 4-bit
- Qwen3-TTS-12Hz-1.7B-VoiceDesignAlibaba1.9B≈1.2 GB at 4-bit
- Hy-MT2-1.8BTencent Hunyuan1.8B≈1.2 GB at 4-bit1 also selling it hosted
- Flux.1-dev-Controlnet-UpscalerJasper.ai1.8B≈1.2 GB at 4-bit
- DeepScaleR-1.5B-PreviewAgentica1.8B≈1.2 GB at 4-bit
- DeepSeek-R1-Distill-Qwen-1.5BDeepSeek1.8B≈1.2 GB at 4-bit3 also selling it hosted
- Nemotron-Research-Reasoning-Qwen-1.5BNVIDIA1.8B≈1.2 GB at 4-bit
- VibeThinker-1.5BWeiboAI1.8B≈1.2 GB at 4-bit
- gte-Qwen2-1.5B-instructAlibaba-NLP1.8B≈1.2 GB at 4-bit
- Qwen3.6-27B-DFlashZ Lab1.7B≈1.1 GB at 4-bit
- BidirLM-1.7B-EmbeddingBidirLM1.7B≈1.1 GB at 4-bit
- F2LLM-1.7Bcodefuse-ai1.7B≈1.1 GB at 4-bit
- F2LLM-v2-1.7Bcodefuse-ai1.7B≈1.1 GB at 4-bit
- Qwen3-ASR-1.7BAlibaba1.7B≈1.1 GB at 4-bit1 also selling it hosted
- SmolLM-1.7BHuggingFaceTB1.7B≈1.1 GB at 4-bit
- SmolLM2-1.7BHuggingFaceTB1.7B≈1.1 GB at 4-bit
- CogVideoX-2bZ.ai1.7B≈1.1 GB at 4-bit
- stablelm-2-1_6bStability AI1.6B≈1.1 GB at 4-bit
- stablelm-2-zephyr-1_6bStability AI1.6B≈1.1 GB at 4-bit
- Dia-1.6BNari Labs1.6B≈1.0 GB at 4-bit
- CrisperWhispernyra labs1.6B≈1.0 GB at 4-bit
- gpt2-xlOpenAI community1.6B≈1.0 GB at 4-bit
- LFM2.5-VL-1.6BLiquid AI1.6B≈1.0 GB at 4-bit
- tts-1.6b-en_frKyutai1.6B≈1.0 GB at 4-bit
- LFM2-VL-1.6BLiquid AI1.6B≈1.0 GB at 4-bit
- musicgen-melodyfacebook1.6B≈1.0 GB at 4-bit
- Qwen2.5-1.5BAlibaba1.5B≈1.0 GB at 4-bit
- Qwen2.5-1.5B-InstructAlibaba1.5B≈1.0 GB at 4-bit1 also selling it hosted
- ReaderLM-v2Jina AI1.5B≈1.0 GB at 4-bit
- reader-lm-1.5bJina AI1.5B≈1.0 GB at 4-bit
- Whisper Large v3OpenAI1.5B≈1.0 GB at 4-bit2 also selling it hosted
- whisper-largeOpenAI1.5B≈1.0 GB at 4-bit
- whisper-large-v2OpenAI1.5B≈1.0 GB at 4-bit
- stable-video-diffusion-img2vidStability AI1.5B≈1.0 GB at 4-bit
- stable-video-diffusion-img2vid-xtStability AI1.5B≈1.0 GB at 4-bit
- stable-video-diffusion-img2vid-xt-1-1Stability AI1.5B≈1.0 GB at 4-bit
- Hymba-1.5B-InstructNVIDIA1.5B≈1.0 GB at 4-bit
- Arch-Router-1.5Bkatanemo1.5B≈1.0 GB at 4-bit
- Hymba-1.5B-BaseNVIDIA1.5B≈1.0 GB at 4-bit
- HyperCLOVAX-SEED-Text-Instruct-1.5BHyperCLOVA X1.5B≈1.0 GB at 4-bit
- Qwen2-1.5B-InstructAlibaba1.5B≈1.0 GB at 4-bit1 also selling it hosted
- stella_en_1.5B_v5NovaSearch1.5B≈1.0 GB at 4-bit
- LFM2-Audio-1.5BLiquid AI1.5B≈1.0 GB at 4-bit
- LFM2.5-Audio-1.5BLiquid AI1.5B≈1.0 GB at 4-bit
- starvector-1b-im2svgstarvector1.4B≈0.9 GB at 4-bit
- i2vgen-xlali-vilab1.4B≈0.9 GB at 4-bit
- phi-1Microsoft1.4B≈0.9 GB at 4-bit
- phi-1_5Microsoft1.4B≈0.9 GB at 4-bit
- text-to-video-ms-1.7bali-vilab1.4B≈0.9 GB at 4-bit
- Ouro-1.4BByteDance1.4B≈0.9 GB at 4-bit
- SSD-1BSegmind1.3B≈0.9 GB at 4-bit1 also selling it hosted
- MiniCPM-V-4.6OpenBMB1.3B≈0.8 GB at 4-bit
- FastWan-QAD-FP8-1.3BFastVideo1.3B≈0.8 GB at 4-bit1 also selling it hosted
- JanusFlow-1.3Bdeepseek-ai1.3B≈0.8 GB at 4-bit
- Wan2.1-T2V-1.3BWan-AI1.3B≈0.8 GB at 4-bit1 also selling it hosted
- Wan2.1-T2V-1.3B-DiffusersWan-AI1.3B≈0.8 GB at 4-bit
- Wan2.1-VACE-1.3BWan-AI1.3B≈0.8 GB at 4-bit
- deepseek-coder-1.3b-instructdeepseek-ai1.3B≈0.8 GB at 4-bit
- gpt-neo-1.3BEleutherAI1.3B≈0.8 GB at 4-bit
- nllb-200-distilled-1.3Bfacebook1.3B≈0.8 GB at 4-bit
- opt-1.3bfacebook1.3B≈0.8 GB at 4-bit
- EXAONE-4.0-1.2BLGAI-EXAONE1.3B≈0.8 GB at 4-bit
- stable-virtual-cameraStability AI1.3B≈0.8 GB at 4-bit
- controlnet-union-sdxl-1.0xinsir1.3B≈0.8 GB at 4-bit
- controlnet-canny-sdxl-1.0🧨Diffusers1.3B≈0.8 GB at 4-bit
- Llama-OuteTTS-1.0-1BOuteAI1.2B≈0.8 GB at 4-bit
- SpatialLM-Llama-1BManycore Research1.2B≈0.8 GB at 4-bit
- Llama 3.2 1B InstructMeta1.2B≈0.8 GB at 4-bit11 also selling it hosted
- stable-audio-open-1.0Stability AI1.2B≈0.8 GB at 4-bit
- LFM2.5-1.2B-BaseLiquid AI1.2B≈0.8 GB at 4-bit
- LFM2.5-1.2B-JPLiquid AI1.2B≈0.8 GB at 4-bit
- MinerU2.5-Pro-2604-1.2BOpenDataLab1.2B≈0.8 GB at 4-bit
- HRM-Text-1BSapient AI1.2B≈0.8 GB at 4-bit
- LFM2-1.2BLiquid AI1.2B≈0.8 GB at 4-bit
- LFM2.5-1.2B-InstructLiquid AI1.2B≈0.8 GB at 4-bit
- LFM2.5-1.2B-ThinkingLiquid AI1.2B≈0.8 GB at 4-bit
- LightOnOCR-1B-1025LightOn AI1.2B≈0.8 GB at 4-bit
- MinerU2.5-2509-1.2BOpenDataLab1.2B≈0.8 GB at 4-bit
- HunyuanOCRTencent Hunyuan1.1B≈0.7 GB at 4-bit
- DeciCoder-1bDeci AI1.1B≈0.7 GB at 4-bit
- TinyLlama-1.1B-Chat-v1.0TinyLlama1.1B≈0.7 GB at 4-bit
- TinyLlama-1.1B-intermediate-step-1431k-3TTinyLlama1.1B≈0.7 GB at 4-bit
- parakeet-rnnt-1.1bNVIDIA1.1B≈0.7 GB at 4-bit
- MiniCPM5-1BOpenBMB1.1B≈0.7 GB at 4-bit
- vaultgemma-1bGoogle1.0B≈0.7 GB at 4-bit
- VibeVoice-Realtime-0.5BMicrosoft1.0B≈0.7 GB at 4-bit
- LightOnOCR-2-1BLightOn AI1.0B≈0.7 GB at 4-bit
- BidirLM-1B-EmbeddingBidirLM1.0B≈0.7 GB at 4-bit