Datasheet for local AI49 models98 GPUsData read 2026-09-25
No adsNo tracking
Check my GPU →

From the RTX 5080 16 GB to the RTX 5090 32 GB: what changes for local AI

+16 GB of VRAM (16 → 32 GB). Of 49 models, 22 newly run well, 16 get a better file or stop offloading, and 11 stay the same.

16 GB GDDR7 · 256-bit · 960 GB/s · Blackwell32 GB GDDR7 · 512-bit · 1792 GB/s · BlackwellData 2026-09-25
22newly run well
16better file
11no change
+16GB more VRAM

01Newly runs well

ModelOn the RTX 5080 16 GBOn the RTX 5090 32 GBNeeded
Qwen-ImageRuns Q4_K_MRuns well FP823.2 GB
Qwen-Image-Edit (2511)Tight Q3_K_MRuns well FP823.5 GB
Ideogram 4Runs Q4_1Runs well FP820.9 GB
HiDream-I1 (Full)Runs Q5_K_MRuns well FP819.9 GB
HiDream-I1 (Dev)Runs Q5_K_MRuns well FP819.9 GB
HunyuanImage 2.1Runs Q4_K_MRuns well FP820.7 GB
Wan 2.1 T2V 14BRuns Q5_K_MRuns well FP818.6 GB
Wan 2.1 I2V 14B 480PRuns Q4_K_MRuns well FP820.7 GB
Wan 2.1 I2V 14B 720PTight Q3_K_MRuns well FP823.2 GB
Wan 2.1 VACE 14BTight Q3_K_SRuns well Q8_023.5 GB
Wan 2.2 T2V A14BRuns Q5_K_MRuns well FP818.6 GB
Wan 2.2 I2V A14BRuns Q5_K_MRuns well FP818.6 GB
Wan 2.2 Animate 14BTight Q3_K_MRuns well FP822.6 GB
Wan Animate 2 (14B)Tight Q3_K_MRuns well Q8_023.4 GB
Wan 2.2 S2V 14BTight Q2_KRuns well FP821.2 GB
SCAIL-2 (character animation)Tight Q3_K_MRuns well FP823.0 GB
HunyuanVideo (13B, original)Runs Q6_KRuns well 16-bit29.9 GB
LTX-Video 13B (0.9.8)Runs Q6_KRuns well FP820.0 GB
LTX-2 (19B)Tight Q3_K_MRuns well Q8_025.2 GB
LTX-2.3 (22B)Tight Q3_K_MRuns well Q8_027.6 GB
LTX-2.5 (22B)Tight Q2_KRuns well Q8_028.4 GB
MiniMax H3 PrunedTight Q3_K_MRuns well FP826.8 GB

“Runs well” means a 16-bit or 8-bit file fits entirely in VRAM.

02A better file, or less offloading

ModelOn the RTX 5080 16 GBOn the RTX 5090 32 GBNeeded
FLUX.2 [dev]Offload only Q2_KRuns Q6_K30.7 GB
MiniMax H3 (33B)Offload only Q4_K_MRuns Q5_K_M29.7 GB
FLUX.1 [dev]Runs well FP8Runs well 16-bit26.1 GB
FLUX.1 [schnell]Runs well FP8Runs well 16-bit26.1 GB
FLUX.1 Kontext [dev]Runs well FP8Runs well 16-bit26.4 GB
FLUX.1 Krea [dev]Runs well FP8Runs well 16-bit26.1 GB
FLUX.1 Fill [dev]Runs well Q8_0Runs well 16-bit26.4 GB
FLUX.2 [klein] 9BRuns well FP8Runs well 16-bit20.5 GB
Krea 2 (Turbo)Runs well FP8Runs well 16-bit28.9 GB
Qwen-Image 2.1Runs well Q8_0Runs well 16-bit16.5 GB
Boogu-Image (Turbo)Runs well FP8Runs well 16-bit23.2 GB
ERNIE-Image (Turbo)Runs well Q8_0Runs well 16-bit18.4 GB
HiDream-O1-ImageRuns well FP8Runs well 16-bit19.2 GB
Stable Diffusion 3.5 LargeRuns well Q8_0Runs well 16-bit18.8 GB
Chroma1-HDRuns well FP8Runs well 16-bit20.1 GB
HunyuanVideo 1.5Runs well FP8Runs well 16-bit21.0 GB

04Speed and features

Memory bandwidth goes from 960 to 1792 GB/s (×1.87). For models that fit on both cards, bandwidth and compute decide the speed; this ratio is a rough first guide, not a benchmark.