Models that run on 256 GB
1161 models with published weights that fit in 256 GB — a Mac Studio with 256. Counted at four-bit quantisation, weights only, leaving the machine about a third of its memory and a gigabyte for context: room for roughly 274 billion parameters. A mixture of experts is counted in full, because it is held in full even though only a few experts compute. 1,161 in all, a hundred to a page; this is page 6 of 12.
A Mac Studio at the top of its configuration, or a serious machine. Almost everything openly published fits, including the large mixtures of experts. At this point the constraint is no longer whether the model loads but whether it generates fast enough to be worth waiting for.
- Qwen3 8BAlibaba8.2B≈5.3 GB at 4-bit8 also selling it hosted
- MisoTTSMiso Labs8.2B≈5.3 GB at 4-bit
- MiniCPM4.1-8BOpenBMB8.2B≈5.3 GB at 4-bit
- granite-3.0-8b-instructIBM Granite8.2B≈5.3 GB at 4-bit1 also selling it hosted
- Flex.2-previewostris8.2B≈5.3 GB at 4-bit
- Flex.1-alphaostris8.2B≈5.3 GB at 4-bit
- flux.1-lite-8B-alphaFreepik8.2B≈5.3 GB at 4-bit
- Qwen3-VL-Embedding-8BAlibaba8.1B≈5.3 GB at 4-bit
- ERNIE-ImageBaidu8.0B≈5.2 GB at 4-bit
- ERNIE-Image-TurboBaidu8.0B≈5.2 GB at 4-bit
- dolphin-2.9-llama3-8bDolphin8.0B≈5.2 GB at 4-bit
- UserLM-8bMicrosoft8.0B≈5.2 GB at 4-bit
- Llama3-ChatQA-1.5-8BNVIDIA8.0B≈5.2 GB at 4-bit
- Aion-RP 1.0 (8B)AionLabs8.0B≈5.2 GB at 4-bit2 also selling it hosted
- DarkIdol-Llama-3.1-8B-Instruct-1.2-Uncensoredaifeifei7988.0B≈5.2 GB at 4-bit
- KernelLLMfacebook8.0B≈5.2 GB at 4-bit
- Llama 3.1 8B InstructMeta8.0B≈5.2 GB at 4-bit20 also selling it hosted
- Llama-3-8B-Instruct-Gradient-1048kDeepSky8.0B≈5.2 GB at 4-bit
- Llama-3-8B-WebMcGill NLP Group8.0B≈5.2 GB at 4-bit
- Llama-3-RefueledRefuel AI8.0B≈5.2 GB at 4-bit
- Llama-3.1-Nemotron-Nano-8B-v1NVIDIA8.0B≈5.2 GB at 4-bit
- Llama-3.1-SuperNova-LiteArcee AI8.0B≈5.2 GB at 4-bit
- Llama3-8B-Chinese-Chatshenzhi-wang8.0B≈5.2 GB at 4-bit
- Meta-Llama-3.1-8B-Instruct-abliteratedmlabonne8.0B≈5.2 GB at 4-bit
- llama-3-Korean-Bllossom-8BMLP-LAB8.0B≈5.2 GB at 4-bit
- aya-23-8BCohere Labs8.0B≈5.2 GB at 4-bit
- c4ai-command-r7b-12-2024Cohere Labs8.0B≈5.2 GB at 4-bit
- Molmo-7B-D-0924Ai28.0B≈5.2 GB at 4-bit
- LLaDA-8B-InstructGSAI-ML8.0B≈5.2 GB at 4-bit
- Apertus-8B-2509swiss-ai8.0B≈5.2 GB at 4-bit
- Apertus-8B-Instruct-2509swiss-ai8.0B≈5.2 GB at 4-bit1 also selling it hosted
- Apertus-v1.5-8Bswiss-ai8.0B≈5.2 GB at 4-bit1 also selling it hosted
- Bio-Medical-MultiModal-Llama-3-8B-V1ContactDoctor8.0B≈5.2 GB at 4-bit
- DeepSeek R1 0528 Qwen3 8BDeepSeek8.0B≈5.2 GB at 4-bit1 also selling it hosted
- DeepSeek-R1-Distill-Llama-8BDeepSeek8.0B≈5.2 GB at 4-bit3 also selling it hosted
- F2LLM-v2-8Bcodefuse-ai8.0B≈5.2 GB at 4-bit
- Foundation-Sec-8Bfdtn-ai8.0B≈5.2 GB at 4-bit
- Hermes 2 Pro Llama 3 8BNous Research8.0B≈5.2 GB at 4-bit2 also selling it hosted
- Idefics3-8B-Llama3HuggingFaceM48.0B≈5.2 GB at 4-bit
- L3 8B Stheno V3.2Sao10K8.0B≈5.2 GB at 4-bit4 also selling it hosted
- L3-8B-Lunaris-v1Sao10K8.0B≈5.2 GB at 4-bit2 also selling it hosted
- Llama-3-8B-Lexi-UncensoredOrenguteng8.0B≈5.2 GB at 4-bit
- Llama-3-ELYZA-JP-8Belyza8.0B≈5.2 GB at 4-bit
- Llama-3-Groq-8B-Tool-UseGroq8.0B≈5.2 GB at 4-bit
- Llama-3-Open-Ko-8Bbeomi8.0B≈5.2 GB at 4-bit
- Llama-3.1-8B-Lexi-Uncensored-V2Orenguteng8.0B≈5.2 GB at 4-bit1 also selling it hosted
- Llama-3.1-Nemotron-Nano-VL-8B-V1NVIDIA8.0B≈5.2 GB at 4-bit
- Llama-3.1-Storm-8Bakjindal532448.0B≈5.2 GB at 4-bit
- Llama-3.1-Tulu-3-8BAllen Institute for AI (Ai2)8.0B≈5.2 GB at 4-bit
- Llama-Guard-3-8BMeta8.0B≈5.2 GB at 4-bit4 also selling it hosted
- Llama3-OpenBioLLM-8Baaditya8.0B≈5.2 GB at 4-bit
- Llama3-TAIDE-LX-8B-Chat-Alpha1taide8.0B≈5.2 GB at 4-bit
- Meta-Llama-Guard-2-8BMeta Llama8.0B≈5.2 GB at 4-bit1 also selling it hosted
- MiniCPM4-8BOpenBMB8.0B≈5.2 GB at 4-bit
- Mistral-NeMo-Minitron-8B-BaseNVIDIA8.0B≈5.2 GB at 4-bit
- NeuralDaredevil-8B-abliteratedmlabonne8.0B≈5.2 GB at 4-bit
- Octen-Embedding-8BOcten8.0B≈5.2 GB at 4-bit
- Qwen3 8B (embeddings)Alibaba8.0B≈5.2 GB at 4-bit6 also selling it hosted
- Qwen3-Reranker-8BAlibaba8.0B≈5.2 GB at 4-bit2 also selling it hosted
- Seed-Coder-8B-ReasoningByteDance Seed8.0B≈5.2 GB at 4-bit
- SenseNova-U1-8B-MoTsensenova8.0B≈5.2 GB at 4-bit
- WeDLM-8B-InstructTencent Hunyuan8.0B≈5.2 GB at 4-bit
- aya-vision-8bCohere Labs8.0B≈5.2 GB at 4-bit
- deepthought-8b-llama-v0.01-alpharuliad8.0B≈5.2 GB at 4-bit
- granite-3.1-8b-instructIBM Granite8.0B≈5.2 GB at 4-bit
- granite-3.3-8b-instructIBM Granite8.0B≈5.2 GB at 4-bit1 also selling it hosted
- granite-speech-3.3-8bIBM Granite8.0B≈5.2 GB at 4-bit
- llama-3-sqlcoder-8bdefog8.0B≈5.2 GB at 4-bit
- llama-embed-nemotron-8bNVIDIA8.0B≈5.2 GB at 4-bit
- Gemma-4-E4BGoogle8.0B≈5.2 GB at 4-bit
- Ling-3.0-tinyInclusionAI7.9B≈5.1 GB at 4-bit
- NV-Embed-v2NVIDIA7.9B≈5.1 GB at 4-bit
- EXAONE-3.0-7.8B-InstructLG AI Research7.8B≈5.1 GB at 4-bit
- phixtral-4x2_8mlabonne7.8B≈5.1 GB at 4-bit
- EXAONE-3.5-7.8B-InstructLGAI-EXAONE7.8B≈5.1 GB at 4-bit
- OpenCoder-8B-Instructinftech.ai7.8B≈5.1 GB at 4-bit
- internlm2_5-7b-chatinternlm7.7B≈5.0 GB at 4-bit
- Qwen-7B-ChatAlibaba7.7B≈5.0 GB at 4-bit
- Qwen1.5-7B-ChatAlibaba7.7B≈5.0 GB at 4-bit
- DeepHat-V1-7BDeepHat7.6B≈5.0 GB at 4-bit
- Marco-o1ATH-MaaS7.6B≈5.0 GB at 4-bit
- Qwen2.5 7B InstructAlibaba7.6B≈5.0 GB at 4-bit8 also selling it hosted
- Qwen2.5-7B-Instruct-1MAlibaba7.6B≈5.0 GB at 4-bit
- VulnLLM-R-7BVirtueAI7.6B≈5.0 GB at 4-bit
- falcon-H1R-7BTechnology Innovation Institute7.6B≈4.9 GB at 4-bit
- llava-v1.6-mistral-7bliuhaotian7.6B≈4.9 GB at 4-bit
- starvector-8b-im2svgstarvector7.5B≈4.9 GB at 4-bit
- Ovis-Image-7BATH-MaaS7.4B≈4.8 GB at 4-bit
- falcon-mamba-7bTechnology Innovation Institute7.3B≈4.7 GB at 4-bit
- CodeQwen1.5-7B-ChatAlibaba7.3B≈4.7 GB at 4-bit
- Ruyi-Mini-7BIamCreateAI7.2B≈4.7 GB at 4-bit
- Starling-LM-7B-alphaBerkeley-Nest7.2B≈4.7 GB at 4-bit
- Starling-LM-7B-betaNexusflow7.2B≈4.7 GB at 4-bit
- dolphin-2.1-mistral-7bDolphin7.2B≈4.7 GB at 4-bit
- dolphin-2.2.1-mistral-7bDolphin7.2B≈4.7 GB at 4-bit
- dolphin-2.8-mistral-7b-v02Dolphin7.2B≈4.7 GB at 4-bit
- openchat-3.5-0106openchat7.2B≈4.7 GB at 4-bit
- Mistral-7B-v0.1Mistral AI7.2B≈4.7 GB at 4-bit
- neural-chat-7b-v3-1Intel7.2B≈4.7 GB at 4-bit
- zephyr-7b-alphaHugging Face H47.2B≈4.7 GB at 4-bit