Best AI Models for 48 GB VRAM
Professional — large 70B models at higher quality quants. Here are all 90 models you can run locally with 48 GB of memory.
Hardware with 48 GB Memory
Runs Comfortably (82)
These models fit with room to spare for context window and OS overhead.
Tight Fit (8)
These models run but with limited context window. Close other apps to free memory.
| Model | Params | Quantization | VRAM | Quality |
|---|---|---|---|---|
| Qwen 2.5 72B | 72B | Q4_K_M | 44.7 GB | 4 |
| Qwen 2.5 VL 72B | 72B | Q4_K_M | 41 GB | 4 |
| Cogito 70B | 70B | Q4_K_M | 43 GB | 4 |
| DeepSeek R1 70B | 70B | Q4_K_M | 43.5 GB | 4 |
| Llama 3.1 70B | 70B | Q4_K_M | 43.5 GB | 4 |
| Llama 3.3 70B | 70B | Q4_K_M | 43.5 GB | 4 |
| Qwen 3.6 35B-A3B | 35B | Q8_0 | 44 GB | 5 |
| InternLM 2.5 20B | 20B | F16 | 42 GB | 5 |
Want to check a specific combination?
Use the compatibility checker to see exactly how a model runs on your specific hardware, with performance estimates.
Open Compatibility Checker