Models that run on 36 GB
1031 models with published weights that fit in 36 GB — a MacBook Pro with 36. Counted at four-bit quantisation, weights only, leaving the machine about a third of its memory and a gigabyte for context: room for roughly 37 billion parameters. A mixture of experts is counted in full, because it is held in full even though only a few experts compute. 1,031 in all, a hundred to a page; this is page 9 of 11.
An Apple configuration, and a comfortable one: it holds what 32 GB holds without the machine feeling tight, which in practice means a longer context or a browser you do not have to close first.
- LFM2.5-Audio-1.5BLiquid AI1.5B≈1.0 GB at 4-bit
- starvector-1b-im2svgstarvector1.4B≈0.9 GB at 4-bit
- i2vgen-xlali-vilab1.4B≈0.9 GB at 4-bit
- phi-1Microsoft1.4B≈0.9 GB at 4-bit
- phi-1_5Microsoft1.4B≈0.9 GB at 4-bit
- text-to-video-ms-1.7bali-vilab1.4B≈0.9 GB at 4-bit
- Ouro-1.4BByteDance1.4B≈0.9 GB at 4-bit
- SSD-1BSegmind1.3B≈0.9 GB at 4-bit1 also selling it hosted
- MiniCPM-V-4.6OpenBMB1.3B≈0.8 GB at 4-bit
- FastWan-QAD-FP8-1.3BFastVideo1.3B≈0.8 GB at 4-bit1 also selling it hosted
- JanusFlow-1.3Bdeepseek-ai1.3B≈0.8 GB at 4-bit
- Wan2.1-T2V-1.3BWan-AI1.3B≈0.8 GB at 4-bit1 also selling it hosted
- Wan2.1-T2V-1.3B-DiffusersWan-AI1.3B≈0.8 GB at 4-bit
- Wan2.1-VACE-1.3BWan-AI1.3B≈0.8 GB at 4-bit
- deepseek-coder-1.3b-instructdeepseek-ai1.3B≈0.8 GB at 4-bit
- gpt-neo-1.3BEleutherAI1.3B≈0.8 GB at 4-bit
- nllb-200-distilled-1.3Bfacebook1.3B≈0.8 GB at 4-bit
- opt-1.3bfacebook1.3B≈0.8 GB at 4-bit
- EXAONE-4.0-1.2BLGAI-EXAONE1.3B≈0.8 GB at 4-bit
- stable-virtual-cameraStability AI1.3B≈0.8 GB at 4-bit
- controlnet-union-sdxl-1.0xinsir1.3B≈0.8 GB at 4-bit
- controlnet-canny-sdxl-1.0🧨Diffusers1.3B≈0.8 GB at 4-bit
- Llama-OuteTTS-1.0-1BOuteAI1.2B≈0.8 GB at 4-bit
- SpatialLM-Llama-1BManycore Research1.2B≈0.8 GB at 4-bit
- Llama 3.2 1B InstructMeta1.2B≈0.8 GB at 4-bit11 also selling it hosted
- stable-audio-open-1.0Stability AI1.2B≈0.8 GB at 4-bit
- LFM2.5-1.2B-BaseLiquid AI1.2B≈0.8 GB at 4-bit
- LFM2.5-1.2B-JPLiquid AI1.2B≈0.8 GB at 4-bit
- MinerU2.5-Pro-2604-1.2BOpenDataLab1.2B≈0.8 GB at 4-bit
- HRM-Text-1BSapient AI1.2B≈0.8 GB at 4-bit
- LFM2-1.2BLiquid AI1.2B≈0.8 GB at 4-bit
- LFM2.5-1.2B-InstructLiquid AI1.2B≈0.8 GB at 4-bit
- LFM2.5-1.2B-ThinkingLiquid AI1.2B≈0.8 GB at 4-bit
- LightOnOCR-1B-1025LightOn AI1.2B≈0.8 GB at 4-bit
- MinerU2.5-2509-1.2BOpenDataLab1.2B≈0.8 GB at 4-bit
- HunyuanOCRTencent Hunyuan1.1B≈0.7 GB at 4-bit
- DeciCoder-1bDeci AI1.1B≈0.7 GB at 4-bit
- TinyLlama-1.1B-Chat-v1.0TinyLlama1.1B≈0.7 GB at 4-bit
- TinyLlama-1.1B-intermediate-step-1431k-3TTinyLlama1.1B≈0.7 GB at 4-bit
- parakeet-rnnt-1.1bNVIDIA1.1B≈0.7 GB at 4-bit
- MiniCPM5-1BOpenBMB1.1B≈0.7 GB at 4-bit
- vaultgemma-1bGoogle1.0B≈0.7 GB at 4-bit
- VibeVoice-Realtime-0.5BMicrosoft1.0B≈0.7 GB at 4-bit
- LightOnOCR-2-1BLightOn AI1.0B≈0.7 GB at 4-bit
- BidirLM-1B-EmbeddingBidirLM1.0B≈0.7 GB at 4-bit
- Gemma3-1B-ITLiteRT Community (FKA TFLite)1.0B≈0.7 GB at 4-bit
- Isaac 0.2 1BPerceptron1.0B≈0.7 GB at 4-bit1 also selling it hosted
- Janus-Pro-1BDeepSeek1.0B≈0.7 GB at 4-bit1 also selling it hosted
- MiniCPM5-1B-Claude-Opus-Fable5-ThinkingGnLOLot1.0B≈0.7 GB at 4-bit
- MolmoE-1B-0924Allen Institute for AI (Ai2)1.0B≈0.7 GB at 4-bit
- Nemotron-3-Embed-1B-BF16NVIDIA1.0B≈0.7 GB at 4-bit1 also selling it hosted
- Nemotron-3-Embed-1B-NVFP4NVIDIA1.0B≈0.7 GB at 4-bit1 also selling it hosted
- Taiyi-Stable-Diffusion-1B-Chinese-v0.1Fengshenbang-LM1.0B≈0.7 GB at 4-bit
- antares-1bfdtn-ai1.0B≈0.7 GB at 4-bit
- canary-1bNVIDIA1.0B≈0.7 GB at 4-bit
- canary-1b-flashNVIDIA1.0B≈0.7 GB at 4-bit
- csm-1bsesame1.0B≈0.7 GB at 4-bit1 also selling it hosted
- granite-4.0-1b-speechIBM watsonx.ai1.0B≈0.7 GB at 4-bit
- granite-4.0-h-1bIBM Granite1.0B≈0.7 GB at 4-bit
- llama-nemotron-embed-vl-1b-v2NVIDIA1.0B≈0.7 GB at 4-bit1 also selling it hosted
- llama-nemotron-rerank-vl-1b-v2NVIDIA1.0B≈0.7 GB at 4-bit1 also selling it hosted
- metavoice-1B-v0.1MetaVoice1.0B≈0.7 GB at 4-bit
- gemma-3-1b-itGoogle1.0B≈0.6 GB at 4-bit
- gemma-3-1b-ptGoogle1.0B≈0.6 GB at 4-bit
- CLIP-ViT-H-14-laion2B-s32B-b79KLAION eV1.0B≈0.6 GB at 4-bit
- canary-1b-v2NVIDIA1.0B≈0.6 GB at 4-bit
- mms-1b-allfacebook1.0B≈0.6 GB at 4-bit
- PaddleOCR-VL-1.5paddlepaddle1.0B≈0.6 GB at 4-bit
- PaddleOCR-VL-1.6paddlepaddle1.0B≈0.6 GB at 4-bit
- MobileLLM-R1-950MAI at Meta0.9B≈0.6 GB at 4-bit
- indic-parler-ttsAI4Bharat0.9B≈0.6 GB at 4-bit
- Qwen3-TTS-12Hz-0.6B-CustomVoiceAlibaba0.9B≈0.6 GB at 4-bit
- PaddleOCR-VL-0.9Bpaddlepaddle0.9B≈0.6 GB at 4-bit1 also selling it hosted
- siglip-so400m-patch14-384Google0.9B≈0.6 GB at 4-bit
- waifu-diffusionhakurei0.9B≈0.6 GB at 4-bit
- LCM_Dreamshaper_v7SimianLuo0.9B≈0.6 GB at 4-bit
- instruct-pix2pixtimbrooks0.9B≈0.6 GB at 4-bit
- Cyberpunk-Anime-DiffusionDGSpitzer0.9B≈0.6 GB at 4-bit
- Dungeons-and-Diffusion0xJustin0.9B≈0.6 GB at 4-bit
- EimisAnimeDiffusion_1.0veimiss0.9B≈0.6 GB at 4-bit
- Ghibli-Diffusionnitrosocke0.9B≈0.6 GB at 4-bit
- GuoFeng3xiaolxl0.9B≈0.6 GB at 4-bit
- Nitro-Diffusionnitrosocke0.9B≈0.6 GB at 4-bit
- Stable_Diffusion_PaperCut_ModelFictiverse0.9B≈0.6 GB at 4-bit
- anything-v5Stable Diffusion API0.9B≈0.6 GB at 4-bit
- dreamlike-anime-1.0dreamlike-art0.9B≈0.6 GB at 4-bit
- dreamlike-diffusion-1.0dreamlike-art0.9B≈0.6 GB at 4-bit
- dreamlike-photoreal-2.0dreamlike-art0.9B≈0.6 GB at 4-bit
- openjourney-v4prompthero0.9B≈0.6 GB at 4-bit
- OvisOCR2ATH-MaaS0.9B≈0.6 GB at 4-bit
- bitnet-b1.58-2B-4TMicrosoft0.8B≈0.6 GB at 4-bit
- gpt2-largeOpenAI community0.8B≈0.5 GB at 4-bit
- VoxCPM1.5OpenBMB0.8B≈0.5 GB at 4-bit
- Qwen3.5-0.8BAlibaba0.8B≈0.5 GB at 4-bit1 also selling it hosted
- t5gemma-2-270m-270mGoogle0.8B≈0.5 GB at 4-bit
- Florence-2-largeMicrosoft0.8B≈0.5 GB at 4-bit
- Florence-2-large-ftMicrosoft0.8B≈0.5 GB at 4-bit
- FastVLM-0.5BApple0.8B≈0.5 GB at 4-bit
- distil-large-v3Whisper Distillation0.8B≈0.5 GB at 4-bit
- distil-large-v2Whisper Distillation0.8B≈0.5 GB at 4-bit