Models that run on 256 GB
1161 models with published weights that fit in 256 GB — a Mac Studio with 256. Counted at four-bit quantisation, weights only, leaving the machine about a third of its memory and a gigabyte for context: room for roughly 274 billion parameters. A mixture of experts is counted in full, because it is held in full even though only a few experts compute. 1,161 in all, a hundred to a page; this is page 3 of 12.
A Mac Studio at the top of its configuration, or a serious machine. Almost everything openly published fits, including the large mixtures of experts. At this point the constraint is no longer whether the model loads but whether it generates fast enough to be worth waiting for.
- Qwen-SEA-LION-v4-32B-ITaisingapore32.0B≈20.8 GB at 4-bit1 also selling it hosted
- Qwen2.5-32BAlibaba32.0B≈20.8 GB at 4-bit1 also selling it hosted
- Qwen2.5-VL-32B-InstructAlibaba32.0B≈20.8 GB at 4-bit3 also selling it hosted
- QwenLong-L1-32BTongyi-Zhiwen32.0B≈20.8 GB at 4-bit
- Sky-T1-32B-PreviewNovaSky-AI32.0B≈20.8 GB at 4-bit1 also selling it hosted
- UIGEN-X-32B-0727Tesslate32.0B≈20.8 GB at 4-bit
- s1-32Bsimplescaling32.0B≈20.8 GB at 4-bit
- xLAM-2-32b-fc-rSalesforce32.0B≈20.8 GB at 4-bit1 also selling it hosted
- Qwen3-Omni-30B-A3B-CaptionerAlibaba31.7B≈20.6 GB at 4-bit
- NVIDIA Nemotron 3.5 Lightning 30B A3BNVIDIA31.6B≈20.5 GB at 4-bit2 also selling it hosted
- Nemotron 3 Nano 30B A3BNVIDIA31.6B≈20.5 GB at 4-bit10 also selling it hosted
- Nemotron-Cascade-2-30B-A3BNVIDIA31.6B≈20.5 GB at 4-bit
- phonellm-alpha-1Pipecat31.6B≈20.5 GB at 4-bit
- Gemma 4 31BGoogle31.3B≈20.3 GB at 4-bit15 also selling it hosted
- Qwen3 VL 30B A3B InstructAlibaba31.1B≈20.2 GB at 4-bit8 also selling it hosted
- Qwen3 VL 30B A3B ThinkingAlibaba31.1B≈20.2 GB at 4-bit7 also selling it hosted
- Gemma-4-31B-it-assistantGoogle31.0B≈20.2 GB at 4-bit
- Gemma-4-31B-it-pearlpearl-ai31.0B≈20.2 GB at 4-bit
- Qwen3-Coder-30B-A3B-Instruct-FP8Alibaba30.5B≈19.8 GB at 4-bit
- MiroThinker-v1.5-30BMiroMind AI30.5B≈19.8 GB at 4-bit
- Qwen3 30B A3BAlibaba30.5B≈19.8 GB at 4-bit11 also selling it hosted
- Qwen3 30B A3B Instruct 2507Alibaba30.5B≈19.8 GB at 4-bit8 also selling it hosted
- Qwen3 30B A3B Thinking 2507Alibaba30.5B≈19.8 GB at 4-bit4 also selling it hosted
- Qwen3 Coder 30B A3B InstructAlibaba30.5B≈19.8 GB at 4-bit11 also selling it hosted
- Tongyi-DeepResearch-30B-A3BAlibaba-NLP30.5B≈19.8 GB at 4-bit
- Hy-MT2-30B-A3BTencent Hunyuan30.0B≈19.5 GB at 4-bit1 also selling it hosted
- Nemotron-Labs-Audex-30B-A3BNVIDIA30.0B≈19.5 GB at 4-bit
- Ovis2.6-30B-A3BATH-MaaS30.0B≈19.5 GB at 4-bit
- Qwen3 Omni 30B A3B InstructAlibaba30.0B≈19.5 GB at 4-bit2 also selling it hosted
- Qwen3 Omni 30B A3B ThinkingAlibaba30.0B≈19.5 GB at 4-bit2 also selling it hosted
- QwenLong-L1.5-30B-A3BTongyi-Zhiwen30.0B≈19.5 GB at 4-bit
- TildeOpen-30bTildeAI30.0B≈19.5 GB at 4-bit
- Wizard-Vicuna-30B-UncensoredQuixi AI30.0B≈19.5 GB at 4-bit
- WizardLM-30B-UncensoredQuixi AI30.0B≈19.5 GB at 4-bit
- granite-4.1-30bIBM Granite30.0B≈19.5 GB at 4-bit
- Muse Glimmer 30BMeta29.8B≈19.4 GB at 4-bit8 also selling it hosted
- stepvideo-t2vStepFun29.3B≈19.1 GB at 4-bit
- medgemma-27b-itGoogle28.8B≈18.7 GB at 4-bit
- translategemma-27b-itGoogle28.8B≈18.7 GB at 4-bit
- ERNIE 4.5 VL 28B A3BBaidu28.0B≈18.2 GB at 4-bit2 also selling it hosted
- ERNIE-4.5-VL-28B-A3B-ThinkingBaidu28.0B≈18.2 GB at 4-bit1 also selling it hosted
- Huihui-Qwen3.8-27B-abliteratedhuihui-ai27.8B≈18.1 GB at 4-bit
- Qwen3.5-27BAlibaba27.8B≈18.1 GB at 4-bit8 also selling it hosted
- Qwen3.5-27B-Claude-4.6-Opus-Reasoning-DistilledJackrong27.8B≈18.1 GB at 4-bit
- Qwen3.6 27BAlibaba27.8B≈18.1 GB at 4-bit12 also selling it hosted
- Qwen3.8 27BAlibaba27.8B≈18.1 GB at 4-bit12 also selling it hosted
- Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-BF16AEON-727.8B≈18.1 GB at 4-bit
- Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAUDavidAU27.8B≈18.1 GB at 4-bit
- Qwen3.8-27B-UncensoredOrcaRouter27.8B≈18.1 GB at 4-bit
- Swift-Qwen3.8-27bukisai27.8B≈18.1 GB at 4-bit
- deepseek-vl2DeepSeek27.5B≈17.9 GB at 4-bit
- Gemma 3 27BGoogle27.4B≈17.8 GB at 4-bit11 also selling it hosted
- Gemma-3-R1984-27BVIDraft27.4B≈17.8 GB at 4-bit
- gemma-3-27b-it-abliteratedmlabonne27.4B≈17.8 GB at 4-bit
- Gemma 2 27BGoogle27.2B≈17.7 GB at 4-bit4 also selling it hosted
- datagemma-rag-27b-itGoogle27.2B≈17.7 GB at 4-bit
- medgemma-27b-text-itGoogle27.0B≈17.6 GB at 4-bit
- C2S-Scale-Gemma-2-27Bvandijklab27.0B≈17.6 GB at 4-bit
- Fara1.5-27BMicrosoft27.0B≈17.6 GB at 4-bit
- Gemma-SEA-LION-v4-27B-ITaisingapore27.0B≈17.6 GB at 4-bit2 also selling it hosted
- Huihui-Qwen3.5-27B-abliteratedhuihui-ai27.0B≈17.6 GB at 4-bit
- Qwen3.6-27B-AEON-Ultimate-Uncensored-BF16AEON-727.0B≈17.6 GB at 4-bit
- Qwen3.8-27B-DFlash2Z Lab27.0B≈17.6 GB at 4-bit
- Qwythos-27B-v1empero-ai27.0B≈17.6 GB at 4-bit
- harrier-oss-v1-27bMicrosoft27.0B≈17.6 GB at 4-bit
- thinkingcap-qwen3.6-27bsference27.0B≈17.6 GB at 4-bit1 also selling it hosted
- Gemma 4 26B A4B ITGoogle26.0B≈16.9 GB at 4-bit
- Gemma-4-26B-A4B-it-assistantGoogle26.0B≈16.9 GB at 4-bit
- diffusiongemma-26B-A4B-itGoogle25.8B≈16.8 GB at 4-bit
- Gemma 4 26B A4BGoogle25.8B≈16.8 GB at 4-bit7 also selling it hosted
- AriaRhymes.AI25.3B≈16.4 GB at 4-bit
- Voxtral Small 24B 2507Mistral AI24.3B≈15.8 GB at 4-bit4 also selling it hosted
- Dolphin Mistral 24B Venice EditionCognitive Computations24.0B≈15.6 GB at 4-bit1 also selling it hosted
- Mistral Small 3.1 24B Instruct 2503Mistral AI24.0B≈15.6 GB at 4-bit3 also selling it hosted
- Mistral-Small-3.2-24B-Instruct-2506Mistral AI24.0B≈15.6 GB at 4-bit1 also selling it hosted
- Mistral Small 3Mistral AI23.6B≈15.3 GB at 4-bit4 also selling it hosted
- sarvam-mSarvam AI23.6B≈15.3 GB at 4-bit
- solar-pro-preview-instructUpstage22.1B≈14.4 GB at 4-bit
- gpt-oss-safeguard-20bOpenAI21.5B≈14.0 GB at 4-bit6 also selling it hosted
- ERNIE-4.5-21B-A3B-PTBaidu21.0B≈13.7 GB at 4-bit1 also selling it hosted
- ERNIE-4.5-21B-A3B-ThinkingBaidu21.0B≈13.7 GB at 4-bit2 also selling it hosted
- GPT OSS 20BOpenAI20.9B≈13.6 GB at 4-bit20 also selling it hosted
- context-1chroma20.9B≈13.6 GB at 4-bit
- gpt-neox-20bEleutherAI20.7B≈13.5 GB at 4-bit
- FireRed-Image-Edit-1.0FireRed Team20.4B≈13.3 GB at 4-bit
- Qwen ImageAlibaba20.4B≈13.3 GB at 4-bit1 also selling it hosted
- Qwen-Image-2512Alibaba20.4B≈13.3 GB at 4-bit
- Qwen-Image-Edit-2509Alibaba20.4B≈13.3 GB at 4-bit
- Qwen-Image-Edit-2511Alibaba20.4B≈13.3 GB at 4-bit
- maple-previewdeepgrove20.2B≈13.1 GB at 4-bit
- Gemma-4-31B-JANG_4M-CRACKdealignai20.2B≈13.1 GB at 4-bit
- cogvlm2-llama3-chat-19Bzai-org19.5B≈12.7 GB at 4-bit
- LTX-2LTX.io18.9B≈12.3 GB at 4-bit
- cogvlm-chat-hfzai-org17.6B≈11.5 GB at 4-bit
- SenseNova-U1.5-8B-MoTsensenova17.5B≈11.4 GB at 4-bit
- Wan2.1-VACE-14BWan-AI17.3B≈11.3 GB at 4-bit
- HiDream-E1-1HiDream.ai17.1B≈11.1 GB at 4-bit
- HiDream-E1-FullHiDream.ai17.1B≈11.1 GB at 4-bit
- HiDream-I1-FullHiDream.ai17.1B≈11.1 GB at 4-bit
- Kimi-VL-A3B-ThinkingMoonshot AI16.4B≈10.7 GB at 4-bit