NVIDIA
GeForce RTX 5090
- Usable fits
- 54 / 62
- Assumed usable
- 29.4 GiB
- Bandwidth
- 1792 GB/s
- Largest fit here
- 46.7B
Largest by parameter count: Mixtral 8x7B Instruct at Q4_K_M.
See all 62 models →Lab · LLM VRAM · GPU lookup
ComputedChoose an accelerator to see every one of the 62 rostered models ranked by on-device fit. A usable fit means the weights plus at least 4k tokens of context fit together, capped at the model’s own trained window; tight and no-fit cases stay visible rather than disappearing.
Capacity and peak bandwidth are published device specifications. “Usable” reserves 8% of dedicated VRAM or 25% of unified memory for the runtime, OS, display, and workspace. Each result page prints both the raw spec and that explicit assumption.
NVIDIA
Largest by parameter count: Mixtral 8x7B Instruct at Q4_K_M.
See all 62 models →NVIDIA
Largest by parameter count: Mixtral 8x7B Instruct at Q3_K_M.
See all 62 models →Largest by parameter count: Mixtral 8x7B Instruct at Q3_K_M.
See all 62 models →NVIDIA
Largest by parameter count: Mixtral 8x7B Instruct at Q3_K_M.
See all 62 models →NVIDIA
Largest by parameter count: Qwen3 30B-A3B at Q3_K_M. 1 more only fit below the usable-context floor.
See all 62 models →NVIDIA
Largest by parameter count: Qwen3 30B-A3B at Q3_K_M. 1 more only fit below the usable-context floor.
See all 62 models →NVIDIA
Largest by parameter count: Qwen3 30B-A3B at Q3_K_M. 1 more only fit below the usable-context floor.
See all 62 models →Largest by parameter count: Qwen3 30B-A3B at Q3_K_M. 1 more only fit below the usable-context floor.
See all 62 models →NVIDIA
Largest by parameter count: Qwen3 30B-A3B at Q3_K_M. 1 more only fit below the usable-context floor.
See all 62 models →NVIDIA
Largest by parameter count: gpt-oss 20B at Q5_K_M. 2 more only fit below the usable-context floor.
See all 62 models →NVIDIA
Largest by parameter count: Qwen2.5 72B Instruct at Q3_K_M.
See all 62 models →NVIDIA
Largest by parameter count: Hy3 at Q3_K_M.
See all 62 models →NVIDIA
Largest by parameter count: Laguna S 2.1 at Q4_K_M.
See all 62 models →NVIDIA
Largest by parameter count: Laguna S 2.1 at Q4_K_M.
See all 62 models →NVIDIA
Largest by parameter count: Qwen2.5 72B Instruct at Q3_K_M.
See all 62 models →NVIDIA
Largest by parameter count: Qwen2.5 72B Instruct at Q3_K_M.
See all 62 models →Largest by parameter count: Hy3 at Q8_0.
See all 62 models →Apple
Largest by parameter count: Laguna S 2.1 at Q6_K.
See all 62 models →Apple
Largest by parameter count: DeepSeek-R1-Distill-Llama 70B at Q3_K_M. 1 more only fit below the usable-context floor.
See all 62 models →Apple
Largest by parameter count: Hermes 4.3 36B at Q3_K_M.
See all 62 models →NVIDIA
Largest by parameter count: Qwen2.5 72B Instruct at Q4_K_M.
See all 62 models →Largest by parameter count: Qwythos 9B Claude Mythos 5 1M at Q4_K_M. 2 more only fit below the usable-context floor.
See all 62 models →