Models that run on 256 GB
1161 models with published weights that fit in 256 GB — a Mac Studio with 256. Counted at four-bit quantisation, weights only, leaving the machine about a third of its memory and a gigabyte for context: room for roughly 274 billion parameters. A mixture of experts is counted in full, because it is held in full even though only a few experts compute. 1,161 in all, a hundred to a page; this is page 10 of 12.
A Mac Studio at the top of its configuration, or a serious machine. Almost everything openly published fits, including the large mixtures of experts. At this point the constraint is no longer whether the model loads but whether it generates fast enough to be worth waiting for.
- Qwen3-ASR-1.7BAlibaba1.7B≈1.1 GB at 4-bit1 also selling it hosted
- SmolLM-1.7BHuggingFaceTB1.7B≈1.1 GB at 4-bit
- SmolLM2-1.7BHuggingFaceTB1.7B≈1.1 GB at 4-bit
- CogVideoX-2bZ.ai1.7B≈1.1 GB at 4-bit
- stablelm-2-1_6bStability AI1.6B≈1.1 GB at 4-bit
- stablelm-2-zephyr-1_6bStability AI1.6B≈1.1 GB at 4-bit
- Dia-1.6BNari Labs1.6B≈1.0 GB at 4-bit
- CrisperWhispernyra labs1.6B≈1.0 GB at 4-bit
- gpt2-xlOpenAI community1.6B≈1.0 GB at 4-bit
- LFM2.5-VL-1.6BLiquid AI1.6B≈1.0 GB at 4-bit
- tts-1.6b-en_frKyutai1.6B≈1.0 GB at 4-bit
- LFM2-VL-1.6BLiquid AI1.6B≈1.0 GB at 4-bit
- musicgen-melodyfacebook1.6B≈1.0 GB at 4-bit
- Qwen2.5-1.5BAlibaba1.5B≈1.0 GB at 4-bit
- Qwen2.5-1.5B-InstructAlibaba1.5B≈1.0 GB at 4-bit1 also selling it hosted
- ReaderLM-v2Jina AI1.5B≈1.0 GB at 4-bit
- reader-lm-1.5bJina AI1.5B≈1.0 GB at 4-bit
- Whisper Large v3OpenAI1.5B≈1.0 GB at 4-bit2 also selling it hosted
- whisper-largeOpenAI1.5B≈1.0 GB at 4-bit
- whisper-large-v2OpenAI1.5B≈1.0 GB at 4-bit
- stable-video-diffusion-img2vidStability AI1.5B≈1.0 GB at 4-bit
- stable-video-diffusion-img2vid-xtStability AI1.5B≈1.0 GB at 4-bit
- stable-video-diffusion-img2vid-xt-1-1Stability AI1.5B≈1.0 GB at 4-bit
- Hymba-1.5B-InstructNVIDIA1.5B≈1.0 GB at 4-bit
- Arch-Router-1.5Bkatanemo1.5B≈1.0 GB at 4-bit
- Hymba-1.5B-BaseNVIDIA1.5B≈1.0 GB at 4-bit
- HyperCLOVAX-SEED-Text-Instruct-1.5BHyperCLOVA X1.5B≈1.0 GB at 4-bit
- Qwen2-1.5B-InstructAlibaba1.5B≈1.0 GB at 4-bit1 also selling it hosted
- stella_en_1.5B_v5NovaSearch1.5B≈1.0 GB at 4-bit
- LFM2-Audio-1.5BLiquid AI1.5B≈1.0 GB at 4-bit
- LFM2.5-Audio-1.5BLiquid AI1.5B≈1.0 GB at 4-bit
- starvector-1b-im2svgstarvector1.4B≈0.9 GB at 4-bit
- i2vgen-xlali-vilab1.4B≈0.9 GB at 4-bit
- phi-1Microsoft1.4B≈0.9 GB at 4-bit
- phi-1_5Microsoft1.4B≈0.9 GB at 4-bit
- text-to-video-ms-1.7bali-vilab1.4B≈0.9 GB at 4-bit
- Ouro-1.4BByteDance1.4B≈0.9 GB at 4-bit
- SSD-1BSegmind1.3B≈0.9 GB at 4-bit1 also selling it hosted
- MiniCPM-V-4.6OpenBMB1.3B≈0.8 GB at 4-bit
- FastWan-QAD-FP8-1.3BFastVideo1.3B≈0.8 GB at 4-bit1 also selling it hosted
- JanusFlow-1.3Bdeepseek-ai1.3B≈0.8 GB at 4-bit
- Wan2.1-T2V-1.3BWan-AI1.3B≈0.8 GB at 4-bit1 also selling it hosted
- Wan2.1-T2V-1.3B-DiffusersWan-AI1.3B≈0.8 GB at 4-bit
- Wan2.1-VACE-1.3BWan-AI1.3B≈0.8 GB at 4-bit
- deepseek-coder-1.3b-instructdeepseek-ai1.3B≈0.8 GB at 4-bit
- gpt-neo-1.3BEleutherAI1.3B≈0.8 GB at 4-bit
- nllb-200-distilled-1.3Bfacebook1.3B≈0.8 GB at 4-bit
- opt-1.3bfacebook1.3B≈0.8 GB at 4-bit
- EXAONE-4.0-1.2BLGAI-EXAONE1.3B≈0.8 GB at 4-bit
- stable-virtual-cameraStability AI1.3B≈0.8 GB at 4-bit
- controlnet-union-sdxl-1.0xinsir1.3B≈0.8 GB at 4-bit
- controlnet-canny-sdxl-1.0🧨Diffusers1.3B≈0.8 GB at 4-bit
- Llama-OuteTTS-1.0-1BOuteAI1.2B≈0.8 GB at 4-bit
- SpatialLM-Llama-1BManycore Research1.2B≈0.8 GB at 4-bit
- Llama 3.2 1B InstructMeta1.2B≈0.8 GB at 4-bit11 also selling it hosted
- stable-audio-open-1.0Stability AI1.2B≈0.8 GB at 4-bit
- LFM2.5-1.2B-BaseLiquid AI1.2B≈0.8 GB at 4-bit
- LFM2.5-1.2B-JPLiquid AI1.2B≈0.8 GB at 4-bit
- MinerU2.5-Pro-2604-1.2BOpenDataLab1.2B≈0.8 GB at 4-bit
- HRM-Text-1BSapient AI1.2B≈0.8 GB at 4-bit
- LFM2-1.2BLiquid AI1.2B≈0.8 GB at 4-bit
- LFM2.5-1.2B-InstructLiquid AI1.2B≈0.8 GB at 4-bit
- LFM2.5-1.2B-ThinkingLiquid AI1.2B≈0.8 GB at 4-bit
- LightOnOCR-1B-1025LightOn AI1.2B≈0.8 GB at 4-bit
- MinerU2.5-2509-1.2BOpenDataLab1.2B≈0.8 GB at 4-bit
- HunyuanOCRTencent Hunyuan1.1B≈0.7 GB at 4-bit
- DeciCoder-1bDeci AI1.1B≈0.7 GB at 4-bit
- TinyLlama-1.1B-Chat-v1.0TinyLlama1.1B≈0.7 GB at 4-bit
- TinyLlama-1.1B-intermediate-step-1431k-3TTinyLlama1.1B≈0.7 GB at 4-bit
- parakeet-rnnt-1.1bNVIDIA1.1B≈0.7 GB at 4-bit
- MiniCPM5-1BOpenBMB1.1B≈0.7 GB at 4-bit
- vaultgemma-1bGoogle1.0B≈0.7 GB at 4-bit
- VibeVoice-Realtime-0.5BMicrosoft1.0B≈0.7 GB at 4-bit
- LightOnOCR-2-1BLightOn AI1.0B≈0.7 GB at 4-bit
- BidirLM-1B-EmbeddingBidirLM1.0B≈0.7 GB at 4-bit
- Gemma3-1B-ITLiteRT Community (FKA TFLite)1.0B≈0.7 GB at 4-bit
- Isaac 0.2 1BPerceptron1.0B≈0.7 GB at 4-bit1 also selling it hosted
- Janus-Pro-1BDeepSeek1.0B≈0.7 GB at 4-bit1 also selling it hosted
- MiniCPM5-1B-Claude-Opus-Fable5-ThinkingGnLOLot1.0B≈0.7 GB at 4-bit
- MolmoE-1B-0924Allen Institute for AI (Ai2)1.0B≈0.7 GB at 4-bit
- Nemotron-3-Embed-1B-BF16NVIDIA1.0B≈0.7 GB at 4-bit1 also selling it hosted
- Nemotron-3-Embed-1B-NVFP4NVIDIA1.0B≈0.7 GB at 4-bit1 also selling it hosted
- Taiyi-Stable-Diffusion-1B-Chinese-v0.1Fengshenbang-LM1.0B≈0.7 GB at 4-bit
- antares-1bfdtn-ai1.0B≈0.7 GB at 4-bit
- canary-1bNVIDIA1.0B≈0.7 GB at 4-bit
- canary-1b-flashNVIDIA1.0B≈0.7 GB at 4-bit
- csm-1bsesame1.0B≈0.7 GB at 4-bit1 also selling it hosted
- granite-4.0-1b-speechIBM watsonx.ai1.0B≈0.7 GB at 4-bit
- granite-4.0-h-1bIBM Granite1.0B≈0.7 GB at 4-bit
- llama-nemotron-embed-vl-1b-v2NVIDIA1.0B≈0.7 GB at 4-bit1 also selling it hosted
- llama-nemotron-rerank-vl-1b-v2NVIDIA1.0B≈0.7 GB at 4-bit1 also selling it hosted
- metavoice-1B-v0.1MetaVoice1.0B≈0.7 GB at 4-bit
- gemma-3-1b-itGoogle1.0B≈0.6 GB at 4-bit
- gemma-3-1b-ptGoogle1.0B≈0.6 GB at 4-bit
- CLIP-ViT-H-14-laion2B-s32B-b79KLAION eV1.0B≈0.6 GB at 4-bit
- canary-1b-v2NVIDIA1.0B≈0.6 GB at 4-bit
- mms-1b-allfacebook1.0B≈0.6 GB at 4-bit
- PaddleOCR-VL-1.5paddlepaddle1.0B≈0.6 GB at 4-bit
- PaddleOCR-VL-1.6paddlepaddle1.0B≈0.6 GB at 4-bit
- MobileLLM-R1-950MAI at Meta0.9B≈0.6 GB at 4-bit