RTX 3090 24 GB for local AI
Which image and video models run on the RTX 3090 24 GB, which file to download for each, and how much VRAM they need.
VRAM24 GB
MemoryGDDR6X
Bus384-bit
Bandwidth936 GB/s
ArchitectureAmpere
Launched2020-09
Launch price$1,499
Runs well43 of 49
Specs: www.nvidia.com · launch: en.wikipedia.org
01What runs on it
| Model | Verdict | Best file | Size | Needed |
|---|---|---|---|---|
| image models | ||||
| FLUX.1 [dev] | Runs well | Q8_0 | 12.7 GB | 15.0 GB |
| FLUX.1 [schnell] | Runs well | Q8_0 | 12.7 GB | 15.0 GB |
| FLUX.1 Kontext [dev] | Runs well | Q8_0 | 12.7 GB | 15.3 GB |
| FLUX.1 Krea [dev] | Runs well | Q8_0 | 12.7 GB | 15.0 GB |
| FLUX.1 Fill [dev] | Runs well | Q8_0 | 12.7 GB | 15.3 GB |
| FLUX.2 [dev] | Runs | Q4_K_M | 20.1 GB | 23.4 GB |
| FLUX.2 [klein] 9B | Runs well | 16-bit | 18.2 GB | 20.5 GB |
| FLUX.2 [klein] 4B | Runs well | 16-bit | 7.8 GB | 9.6 GB |
| Krea 2 (Turbo) | Runs well | Q8_0 | 13.7 GB | 16.3 GB |
| Qwen-Image | Runs well | FP8 | 20.4 GB | 23.2 GB |
| Qwen-Image-Edit (2511) | Runs well | FP8 | 20.5 GB | 23.5 GB |
| Qwen-Image 2.1 | Runs well | 16-bit | 14.2 GB | 16.5 GB |
| Z-Image Turbo | Runs well | 16-bit | 12.3 GB | 14.3 GB |
| Z-Image (base) | Runs well | 16-bit | 12.3 GB | 14.3 GB |
| Ideogram 4 | Runs well | Q8_0 | 10.1 GB | 22.6 GB |
| Boogu-Image (Turbo) | Runs well | 16-bit | 20.6 GB | 23.2 GB |
| ERNIE-Image (Turbo) | Runs well | 16-bit | 16.1 GB | 18.4 GB |
| HiDream-O1-Image | Runs well | 16-bit | 16.4 GB | 19.2 GB |
| Mage-Flow (Microsoft) | Runs well | 16-bit | 8.2 GB | 10.0 GB |
| Lumina Image 2.0 | Runs well | 16-bit | 5.2 GB | 7.0 GB |
| HiDream-I1 (Full) | Runs well | Q8_0 | 18.7 GB | 21.5 GB |
| HiDream-I1 (Dev) | Runs well | Q8_0 | 18.7 GB | 21.5 GB |
| Stable Diffusion 3.5 Large | Runs well | 16-bit | 16.5 GB | 18.8 GB |
| Stable Diffusion 3.5 Medium | Runs well | 16-bit | 5.1 GB | 6.9 GB |
| Chroma1-HD | Runs well | 16-bit | 17.8 GB | 20.1 GB |
| SDXL 1.0 | Runs well | 16-bit | 6.9 GB | 7.1 GB |
| Illustrious XL / Pony (SDXL anime) | Runs well | 16-bit | 6.9 GB | 7.1 GB |
| Stable Diffusion 1.5 | Runs well | 16-bit | 2.1 GB | 3.3 GB |
| HunyuanImage 2.1 | Runs well | Q8_0 | 19.8 GB | 23.1 GB |
| video models | ||||
| Wan 2.1 T2V 14B | Runs well | Q8_0 | 15.9 GB | 20.2 GB |
| Wan 2.1 T2V 1.3B | Runs well | 16-bit | 2.8 GB | 5.6 GB |
| Wan 2.1 I2V 14B 480P | Runs well | Q8_0 | 18.1 GB | 22.4 GB |
| Wan 2.1 I2V 14B 720P | Runs well | FP8 | 16.4 GB | 23.2 GB |
| Wan 2.1 VACE 14B | Runs well | Q8_0 | 18.7 GB | 23.5 GB |
| Wan 2.2 T2V A14B | Runs well | Q8_0 | 15.4 GB | 19.7 GB |
| Wan 2.2 I2V A14B | Runs well | Q8_0 | 15.4 GB | 19.7 GB |
| Wan 2.2 TI2V 5B | Runs well | 16-bit | 10.0 GB | 13.8 GB |
| Wan 2.2 Animate 14B | Runs well | FP8 | 17.3 GB | 22.6 GB |
| Wan Animate 2 (14B) | Runs well | Q8_0 | 18.1 GB | 23.4 GB |
| Wan 2.2 S2V 14B | Runs well | FP8 | 16.4 GB | 21.2 GB |
| SCAIL-2 (character animation) | Runs well | Q8_0 | 18.1 GB | 23.4 GB |
| HunyuanVideo (13B, original) | Runs well | Q8_0 | 14.0 GB | 18.3 GB |
| LTX-Video 13B (0.9.8) | Runs well | Q8_0 | 14.0 GB | 18.3 GB |
| HunyuanVideo 1.5 | Runs well | 16-bit | 16.7 GB | 21.0 GB |
| LTX-2 (19B) | Runs | Q6_K | 16.0 GB | 20.8 GB |
| LTX-2.3 (22B) | Runs | Q6_K | 17.8 GB | 22.6 GB |
| LTX-2.5 (22B) | Runs | Q6_K | 18.7 GB | 23.5 GB |
| MiniMax H3 (33B) | Tight | Q3_K_M | 15.6 GB | 21.4 GB |
| MiniMax H3 Pruned | Runs | Q6_K | 16.7 GB | 22.5 GB |
Calculated from real file sizes plus working memory. How the numbers work.
02Good to know
The RTX 3090 24 GB (Ampere) has no FP8 compute. FP8 files still load and save memory, but ComfyUI converts them back before the maths, so there is no speed gain — Q8_0 GGUF is the closer-to-original 8-bit choice. ComfyUI can compute INT8 files natively on NVIDIA, which is worth a try where a model offers one.
03Measured and reported results
| Label | Model | Setup | Result | Peak VRAM | Date | Source |
|---|---|---|---|---|---|---|
| reported | FLUX.1 dev | fp8 (ComfyUI template) “Nvidia 3090: 26s” ComfyUI 'GPU Benchmark Flux DEV fp8' thread: stock Flux dev fp8 workflow template, time of 2nd/3rd run (no loading); resolution/steps not stated in thread | 26 s / image | — | 2025-07-22 | github.com → |
| reported | HiDream-I1 | HiDream Full fp8 (T5 fp8, Llama3.1 fp8 scaled) “ComfyUI Full example from their page at https://comfyanonymous.github.io/ComfyUI_examples/hidream/ runs at 115.09s on my RTX 3090, from cold start.” HiDream-I1 Full via ComfyUI example workflow, Windows portable; cold start incl. load; resolution/steps not restated | 115.09 s / image | — | 2025-04-18 | huggingface.co → |
| reported | Illustrious / Pony | Illustrious-XL-v2.0 · 1024x1024 “NVIDIA RTX 3090 | 4.00it/s | 0.25s/it | CUDA 12.9 | ComfyUI (Unknown) | Arch Linux” Community speed list; ComfyUI, Euler/Normal, CFG 8, square 1:1, batch speed only (no per-image time); date = list last-updated date; SDXL 1024px section | 0.25 s/it | — | 2025-11-29 | huggingface.co → |
| reported | MiniMax H3 | transformer variant not stated; NVFP4 text encoder · 832x480 “render 23 min 17 s ... peak VRAM ~19.8 GB of 24 GB” 15.08 s clip at 24 fps with audio; needs --disable-pinned-memory on 32 GB RAM; 23 min 17 s = 1397 s; no date shown | 1397 s / clip (362 frames) | 19.8 GB | — | github.com → |
| reported | SDXL | sd_xl_base_1.0 · 1024x1024 “Prompt executed in 6.16 seconds” ComfyUI 'GPU Benchmark' thread: default workflow, SDXL 1.0 base, 1024x1024, seed 1, second run | 6.16 s / image | — | 2024-03-06 | github.com → |
| reported | Z-Image Turbo | Tongyi-MAI/Z-Image-Turbo · 1024x1024 “NVIDIA RTX 3090 | 0.50it/s | 2.01s/it | CUDA 12.6 | ComfyUI (5151cff) | Arch Linux” Community speed list; ComfyUI, Euler/Normal, CFG 8, square 1:1, batch speed only (no per-image time); date = list last-updated date; Z-Image 1024px section; CFG 8 (above Turbo's us | 2.01 s/it | — | 2025-11-29 | huggingface.co → |
Reported results are other people's numbers, copied as published, with a link. Settings, drivers and ComfyUI versions differ, so compare them with care. Send yours.