RTX 5090 32 GB for local AI
Which image and video models run on the RTX 5090 32 GB, which file to download for each, and how much VRAM they need.
VRAM32 GB
MemoryGDDR7
Bus512-bit
Bandwidth1792 GB/s
ArchitectureBlackwell
Launched2025-01
Launch price$1,999
Runs well47 of 49
Specs: www.nvidia.com · launch: nvidianews.nvidia.com
01What runs on it
| Model | Verdict | Best file | Size | Needed |
|---|---|---|---|---|
| image models | ||||
| FLUX.1 [dev] | Runs well | 16-bit | 23.8 GB | 26.1 GB |
| FLUX.1 [schnell] | Runs well | 16-bit | 23.8 GB | 26.1 GB |
| FLUX.1 Kontext [dev] | Runs well | 16-bit | 23.8 GB | 26.4 GB |
| FLUX.1 Krea [dev] | Runs well | 16-bit | 23.8 GB | 26.1 GB |
| FLUX.1 Fill [dev] | Runs well | 16-bit | 23.8 GB | 26.4 GB |
| FLUX.2 [dev] | Runs | Q6_K | 27.4 GB | 30.7 GB |
| FLUX.2 [klein] 9B | Runs well | 16-bit | 18.2 GB | 20.5 GB |
| FLUX.2 [klein] 4B | Runs well | 16-bit | 7.8 GB | 9.6 GB |
| Krea 2 (Turbo) | Runs well | 16-bit | 26.3 GB | 28.9 GB |
| Qwen-Image | Runs well | FP8 | 20.4 GB | 23.2 GB |
| Qwen-Image-Edit (2511) | Runs well | FP8 | 20.5 GB | 23.5 GB |
| Qwen-Image 2.1 | Runs well | 16-bit | 14.2 GB | 16.5 GB |
| Z-Image Turbo | Runs well | 16-bit | 12.3 GB | 14.3 GB |
| Z-Image (base) | Runs well | 16-bit | 12.3 GB | 14.3 GB |
| Ideogram 4 | Runs well | FP8 | 9.3 GB | 20.9 GB |
| Boogu-Image (Turbo) | Runs well | 16-bit | 20.6 GB | 23.2 GB |
| ERNIE-Image (Turbo) | Runs well | 16-bit | 16.1 GB | 18.4 GB |
| HiDream-O1-Image | Runs well | 16-bit | 16.4 GB | 19.2 GB |
| Mage-Flow (Microsoft) | Runs well | 16-bit | 8.2 GB | 10.0 GB |
| Lumina Image 2.0 | Runs well | 16-bit | 5.2 GB | 7.0 GB |
| HiDream-I1 (Full) | Runs well | FP8 | 17.1 GB | 19.9 GB |
| HiDream-I1 (Dev) | Runs well | FP8 | 17.1 GB | 19.9 GB |
| Stable Diffusion 3.5 Large | Runs well | 16-bit | 16.5 GB | 18.8 GB |
| Stable Diffusion 3.5 Medium | Runs well | 16-bit | 5.1 GB | 6.9 GB |
| Chroma1-HD | Runs well | 16-bit | 17.8 GB | 20.1 GB |
| SDXL 1.0 | Runs well | 16-bit | 6.9 GB | 7.1 GB |
| Illustrious XL / Pony (SDXL anime) | Runs well | 16-bit | 6.9 GB | 7.1 GB |
| Stable Diffusion 1.5 | Runs well | 16-bit | 2.1 GB | 3.3 GB |
| HunyuanImage 2.1 | Runs well | FP8 | 17.4 GB | 20.7 GB |
| video models | ||||
| Wan 2.1 T2V 14B | Runs well | FP8 | 14.3 GB | 18.6 GB |
| Wan 2.1 T2V 1.3B | Runs well | 16-bit | 2.8 GB | 5.6 GB |
| Wan 2.1 I2V 14B 480P | Runs well | FP8 | 16.4 GB | 20.7 GB |
| Wan 2.1 I2V 14B 720P | Runs well | FP8 | 16.4 GB | 23.2 GB |
| Wan 2.1 VACE 14B | Runs well | Q8_0 | 18.7 GB | 23.5 GB |
| Wan 2.2 T2V A14B | Runs well | FP8 | 14.3 GB | 18.6 GB |
| Wan 2.2 I2V A14B | Runs well | FP8 | 14.3 GB | 18.6 GB |
| Wan 2.2 TI2V 5B | Runs well | 16-bit | 10.0 GB | 13.8 GB |
| Wan 2.2 Animate 14B | Runs well | FP8 | 17.3 GB | 22.6 GB |
| Wan Animate 2 (14B) | Runs well | Q8_0 | 18.1 GB | 23.4 GB |
| Wan 2.2 S2V 14B | Runs well | FP8 | 16.4 GB | 21.2 GB |
| SCAIL-2 (character animation) | Runs well | FP8 | 17.7 GB | 23.0 GB |
| HunyuanVideo (13B, original) | Runs well | 16-bit | 25.6 GB | 29.9 GB |
| LTX-Video 13B (0.9.8) | Runs well | FP8 | 15.7 GB | 20.0 GB |
| HunyuanVideo 1.5 | Runs well | 16-bit | 16.7 GB | 21.0 GB |
| LTX-2 (19B) | Runs well | Q8_0 | 20.4 GB | 25.2 GB |
| LTX-2.3 (22B) | Runs well | Q8_0 | 22.8 GB | 27.6 GB |
| LTX-2.5 (22B) | Runs well | Q8_0 | 23.6 GB | 28.4 GB |
| MiniMax H3 (33B) | Runs | Q5_K_M | 23.9 GB | 29.7 GB |
| MiniMax H3 Pruned | Runs well | FP8 | 21.0 GB | 26.8 GB |
Calculated from real file sizes plus working memory. How the numbers work.
02Good to know
The RTX 5090 32 GB is a Blackwell GPU with FP8 and FP4 hardware: ComfyUI computes Comfy-Org's FP8 files natively here, and NVFP4 files (where a model offers them) are faster still.
03Measured and reported results
| Label | Model | Setup | Result | Peak VRAM | Date | Source |
|---|---|---|---|---|---|---|
| reported | FLUX.1 dev | fp8 (ComfyUI template) “Getting 8.78s at 2.38it/s for 3 runs.” ComfyUI 'GPU Benchmark Flux DEV fp8' thread: stock Flux dev fp8 workflow template, time of 2nd/3rd run (no loading); resolution/steps not stated in thread; Inno3D RTX 5090 X3 OC | 8.78 s / image · 2.38 it/s | — | 2025-08-05 | github.com → |
| reported | HunyuanVideo 1.5 | 720p model · 848x480 “I was able to generate 5 seconds video in 284s at 24fps using the 720p model at 848*480” ComfyUI; 5 s video at 24 fps; steps not stated | 284 s / clip | — | 2025-11-22 | huggingface.co → |
| reported | Illustrious / Pony | Illustrious-XL-v2.0 · 1024x1024 “NVIDIA RTX 5090 | 8.95it/s | 0.11s/it | CUDA 12.8 | ComfyUI (Unknown) | Windows 11 24H2” Community speed list; ComfyUI, Euler/Normal, CFG 8, square 1:1, batch speed only (no per-image time); date = list last-updated date; SDXL 1024px section | 0.11 s/it | — | 2025-11-29 | huggingface.co → |
| reported | LTX-2 | LTX-2 19B NVFP4 · 720p “generating 143 frames at 720p takes about 66 seconds end‑to‑end in ComfyUI” Issue says this is ~30-40% slower than NVIDIA's expected 40-45 s; steps not stated | 66 s / clip (143 frames) | — | 2026-01-08 | github.com → |
| reported | MiniMax H3 Pruned | minimax_h3_fl2va_pruned_nvfp4 · 864x480 · 10 steps “175 s for a 864×480 ten-second clip, 26,914 MiB peak VRAM.” ComfyUI 0.30.1, 500 W power cap; 26,914 MiB = 26.28 GiB; INT8-ConvRot = 185 s / 28,581 MiB | 175 s / clip (243 frames) | 26.28 GB | 2026-08-04 | ai-muninn.com → |
| reported | Z-Image Turbo | fp8_e4m3fn (per log) “without SA: 0.95 seg” 'seg' = segundos (seconds) per image; SageAttention variants gave 0.87-0.91 s; resolution/steps not stated | 0.95 s / image | — | 2025-11-27 | github.com → |
Reported results are other people's numbers, copied as published, with a link. Settings, drivers and ComfyUI versions differ, so compare them with care. Send yours.