Models that run on 96 GB
1115 models with published weights that fit in 96 GB — a MacBook Pro with 96. Counted at four-bit quantisation, weights only, leaving the machine about a third of its memory and a gigabyte for context: room for roughly 102 billion parameters. A mixture of experts is counted in full, because it is held in full even though only a few experts compute. 1,115 in all, a hundred to a page; this is page 4 of 12.
An Apple configuration for people who intend to run models rather than occasionally try one. Comfortably holds the seventy-billion class with a very long context, or a mixture-of-experts model whose active parameters are few but whose weights are all resident.
- Llama-2-13b-hfMeta Llama13.0B≈8.5 GB at 4-bit
- Baichuan2-13B-ChatBaichuan Intelligent Technology13.0B≈8.5 GB at 4-bit
- CodeLlama-13b-Instruct-hfCode Llama13.0B≈8.5 GB at 4-bit
- LLaVA-13b-delta-v0liuhaotian13.0B≈8.5 GB at 4-bit
- Llama-2-13bMeta Llama13.0B≈8.5 GB at 4-bit
- Llama-2-13b-chatMeta Llama13.0B≈8.5 GB at 4-bit1 also selling it hosted
- Llama-2-13b-chat-hfMeta13.0B≈8.5 GB at 4-bit1 also selling it hosted
- Llama2-13B-TiefighterKoboldAI13.0B≈8.5 GB at 4-bit1 also selling it hosted
- Llama2-Chinese-13b-ChatFlagAlpha13.0B≈8.5 GB at 4-bit
- MythoMax 13BGryphe13.0B≈8.5 GB at 4-bit4 also selling it hosted
- Nous Hermes Llama2 13B13.0B≈8.5 GB at 4-bit2 also selling it hosted
- OpenOrca-Platypus2-13BOpenOrca13.0B≈8.5 GB at 4-bit
- Orca-2-13bMicrosoft13.0B≈8.5 GB at 4-bit
- ReMM SLERP 13BUndi9513.0B≈8.5 GB at 4-bit1 also selling it hosted
- WhiteRabbitNeo-13B-v1WhiteRabbitNeo13.0B≈8.5 GB at 4-bit
- Wizard-Vicuna-13B-UncensoredQuixi AI13.0B≈8.5 GB at 4-bit
- Wizard-Vicuna-13B-Uncensored-HFTheBloke13.0B≈8.5 GB at 4-bit
- WizardLM-13B-UncensoredQuixi AI13.0B≈8.5 GB at 4-bit
- WizardLM-13B-V1.2WizardLM Team13.0B≈8.5 GB at 4-bit
- Ziya-LLaMA-13B-v1Fengshenbang-LM13.0B≈8.5 GB at 4-bit
- chronos-hermes-13b-v2Austism13.0B≈8.5 GB at 4-bit2 also selling it hosted
- jais-13binception4213.0B≈8.5 GB at 4-bit
- jais-13b-chatinception4213.0B≈8.5 GB at 4-bit1 also selling it hosted
- llama-13bhuggyllama13.0B≈8.5 GB at 4-bit
- mythalion-13bPygmalionAI13.0B≈8.5 GB at 4-bit
- open_llama_13bOpenLM Research13.0B≈8.5 GB at 4-bit
- prometheus-13b-v1.0prometheus-eval13.0B≈8.5 GB at 4-bit
- ruGPT-3.5-13Bai-forever13.0B≈8.5 GB at 4-bit
- stable-vicuna-13b-deltaCarperAI13.0B≈8.5 GB at 4-bit
- vicuna-13b-v1.5Large Model Systems Organization13.0B≈8.5 GB at 4-bit
- vicuna-13b-v1.5-16kLarge Model Systems Organization13.0B≈8.5 GB at 4-bit
- Krea-2-RawKrea12.8B≈8.3 GB at 4-bit
- Krea-2-TurboKrea12.8B≈8.3 GB at 4-bit
- Wayfarer-12BLatitude12.2B≈8.0 GB at 4-bit
- Mistral NemoMistral AI12.2B≈8.0 GB at 4-bit10 also selling it hosted
- Gemma 3 12BGoogle12.2B≈7.9 GB at 4-bit9 also selling it hosted
- Gemma-3-R1984-12BVIDraft12.2B≈7.9 GB at 4-bit
- Mellum2-12B-A2.5B-ThinkingJetBrains12.1B≈7.9 GB at 4-bit
- Llama Guard 4 12BMeta12.0B≈7.8 GB at 4-bit6 also selling it hosted
- Huihui-gemma-4-12B-coder-fable5-composer2.5-v1-abliteratedhuihui-ai12.0B≈7.8 GB at 4-bit
- KaLM-Embedding-Gemma3-12B-2511Tencent Hunyuan12.0B≈7.8 GB at 4-bit
- MN-12B-Celeste-V1.9nothingiisreal12.0B≈7.8 GB at 4-bit
- NVIDIA-Nemotron-Nano-12B-v2NVIDIA12.0B≈7.8 GB at 4-bit2 also selling it hosted
- NemoMix-Unleashed-12BMarinaraSpaghetti12.0B≈7.8 GB at 4-bit
- Pixtral 12B 2409Mistral AI12.0B≈7.8 GB at 4-bit3 also selling it hosted
- gemma-3-12b-it-qat-q4_0-unquantizedGoogle12.0B≈7.8 GB at 4-bit
- oasst-sft-1-pythia-12bOpenAssistant12.0B≈7.8 GB at 4-bit
- oasst-sft-4-pythia-12b-epoch-3.5OpenAssistant12.0B≈7.8 GB at 4-bit
- pythia-12bEleutherAI12.0B≈7.8 GB at 4-bit1 also selling it hosted
- Gemma-4-12BGoogle12.0B≈7.8 GB at 4-bit
- Gemma-4-12B-itGoogle12.0B≈7.8 GB at 4-bit
- Huihui-gemma-4-12B-it-abliteratedhuihui-ai12.0B≈7.8 GB at 4-bit
- FLUX.1-Canny-devBlack Forest Labs11.9B≈7.7 GB at 4-bit
- FLUX.1-Depth-devBlack Forest Labs11.9B≈7.7 GB at 4-bit
- AWPortrait-FLShakker Labs11.9B≈7.7 GB at 4-bit
- FLUX.1-Krea-devBlack Forest Labs11.9B≈7.7 GB at 4-bit
- PixelWave_FLUX.1-dev_03Mikey And Friends11.9B≈7.7 GB at 4-bit
- SRPOTencent Hunyuan11.9B≈7.7 GB at 4-bit
- UltraFlux-v1Owen77711.9B≈7.7 GB at 4-bit
- FLUX.1 [schnell] (Turbo)Black Forest Labs11.9B≈7.7 GB at 4-bit3 also selling it hosted
- OpenFLUX.1ostris11.9B≈7.7 GB at 4-bit
- shuttle-3-diffusionShuttleAI11.9B≈7.7 GB at 4-bit
- falcon-11BTechnology Innovation Institute11.1B≈7.2 GB at 4-bit
- Bielik-11B-v3.0-Instructspeakleash11.0B≈7.2 GB at 4-bit1 also selling it hosted
- Llama 3.2 11B Vision InstructMeta11.0B≈7.2 GB at 4-bit7 also selling it hosted
- Llama Guard 3 11B VisionMeta11.0B≈7.2 GB at 4-bit2 also selling it hosted
- Llama-3.2V-11B-cotXkev11.0B≈7.2 GB at 4-bit
- YanoljaNEXT-EEVE-Instruct-10.8Byanolja10.8B≈7.0 GB at 4-bit
- HyperCLOVAX-SEED-Omni-8BHyperCLOVA X10.7B≈7.0 GB at 4-bit
- Fimbulvetr-11B-v2Sao10K10.7B≈7.0 GB at 4-bit
- Solar-10.7B-Instruct-v1.0Upstage10.7B≈7.0 GB at 4-bit
- Solar-10.7B-v1.0Upstage10.7B≈7.0 GB at 4-bit
- Llama-3.2-11B-VisionMeta Llama10.6B≈6.9 GB at 4-bit
- Boogu-Image-0.1-EditBoogu10.3B≈6.7 GB at 4-bit
- Step3-VL-10BStepFun10.2B≈6.6 GB at 4-bit
- Kimi-Audio-7B-InstructMoonshot AI9.8B≈6.3 GB at 4-bit
- Qwen3.5 9BAlibaba9.7B≈6.3 GB at 4-bit8 also selling it hosted
- Qwen3.8-9B-Distillempero-ai9.7B≈6.3 GB at 4-bit
- liftDatalab9.7B≈6.3 GB at 4-bit
- MiniCPM-SALAOpenBMB9.5B≈6.2 GB at 4-bit
- Fun-Audio-Chat-8BQwenAudio9.5B≈6.1 GB at 4-bit
- OmniCoder-9BTesslate9.4B≈6.1 GB at 4-bit
- Qwen3.5-9B-Claude-4.6-HighIQ-THINKING-HERETIC-UNCENSOREDDavidAU9.4B≈6.1 GB at 4-bit
- Qwythos-9B-Claude-Mythos-5-1Mempero-ai9.4B≈6.1 GB at 4-bit
- fuyu-8bAdept AI Labs9.4B≈6.1 GB at 4-bit
- MiniCPM-o-4_5OpenBMB9.4B≈6.1 GB at 4-bit
- VibeVoice-7BVibeVoice Community (Unofficial)9.3B≈6.1 GB at 4-bit
- VibeVoice-Largeaoi-ot9.3B≈6.1 GB at 4-bit
- kugelaudio-0-openKugelaudio9.3B≈6.1 GB at 4-bit
- ideogram-4-fp8Ideogram9.3B≈6.0 GB at 4-bit
- moondream3-previewmoondream9.3B≈6.0 GB at 4-bit
- bge-multilingual-gemma2BAAI9.2B≈6.0 GB at 4-bit
- gemma-2-9bGoogle9.2B≈6.0 GB at 4-bit
- EuroLLM-9B-InstructUTTER - Unified Transcription and Translation for Extended Reality9.2B≈5.9 GB at 4-bit
- FLUX.2-klein-9b-kvBlack Forest Labs9.1B≈5.9 GB at 4-bit
- AutoGLM-Phone-9B-MultilingualZ.ai9.0B≈5.9 GB at 4-bit3 also selling it hosted
- EuroLLM-9BUTTER - Unified Transcription and Translation for Extended Reality9.0B≈5.9 GB at 4-bit
- FLUX.2-klein-9b-fp8Black Forest Labs9.0B≈5.9 GB at 4-bit
- FLUX.2-klein-base-9BBlack Forest Labs9.0B≈5.9 GB at 4-bit
- FLUX.2-klein-base-9b-fp8Black Forest Labs9.0B≈5.9 GB at 4-bit