Datasheet for local AI49 models98 GPUsData read 2026-09-25
No adsNo tracking
Check my GPU →

From the RTX 3070 8 GB to the RTX 3090 24 GB: what changes for local AI

+16 GB of VRAM (8 → 24 GB). Of 49 models, 35 newly run well, 8 get a better file or stop offloading, and 6 stay the same.

8 GB GDDR6 · 256-bit · 448 GB/s · Ampere24 GB GDDR6X · 384-bit · 936 GB/s · AmpereData 2026-09-25
35newly run well
8better file
6no change
+16GB more VRAM

01Newly runs well

ModelOn the RTX 3070 8 GBOn the RTX 3090 24 GBNeeded
FLUX.1 [dev]Tight Q3_K_SRuns well Q8_015.0 GB
FLUX.1 [schnell]Tight Q3_K_SRuns well Q8_015.0 GB
FLUX.1 Kontext [dev]Tight Q3_K_MRuns well Q8_015.3 GB
FLUX.1 Krea [dev]Tight Q3_K_MRuns well Q8_015.0 GB
FLUX.1 Fill [dev]Tight Q3_K_SRuns well Q8_015.3 GB
FLUX.2 [klein] 9BTight Q3_K_MRuns well 16-bit20.5 GB
Krea 2 (Turbo)Tight Q2_KRuns well Q8_016.3 GB
Qwen-ImageOffload only Q2_KRuns well FP823.2 GB
Qwen-Image-Edit (2511)Offload only Q4_K_MRuns well FP823.5 GB
Qwen-Image 2.1Runs Q5_K_MRuns well 16-bit16.5 GB
Z-Image TurboRuns Q6_KRuns well 16-bit14.3 GB
Z-Image (base)Runs Q5_K_MRuns well 16-bit14.3 GB
Ideogram 4Offload only Q4_1Runs well Q8_022.6 GB
Boogu-Image (Turbo)Offload only Q4_1Runs well 16-bit23.2 GB
ERNIE-Image (Turbo)Runs Q4_K_MRuns well 16-bit18.4 GB
HiDream-O1-ImageOffload only FP8Runs well 16-bit19.2 GB
HiDream-I1 (Full)Offload only Q2_KRuns well Q8_021.5 GB
HiDream-I1 (Dev)Offload only Q2_KRuns well Q8_021.5 GB
Stable Diffusion 3.5 LargeRuns Q4_1Runs well 16-bit18.8 GB
Chroma1-HDRuns Q4_K_MRuns well 16-bit20.1 GB
HunyuanImage 2.1Offload only Q4_K_MRuns well Q8_023.1 GB
Wan 2.1 T2V 14BOffload only Q4_K_MRuns well Q8_020.2 GB
Wan 2.1 I2V 14B 480POffload only Q4_K_MRuns well Q8_022.4 GB
Wan 2.1 I2V 14B 720POffload only Q4_K_MRuns well FP823.2 GB
Wan 2.1 VACE 14BOffload only Q4_K_MRuns well Q8_023.5 GB
Wan 2.2 T2V A14BOffload only Q2_KRuns well Q8_019.7 GB
Wan 2.2 I2V A14BOffload only Q2_KRuns well Q8_019.7 GB
Wan 2.2 TI2V 5BRuns Q5_K_MRuns well 16-bit13.8 GB
Wan 2.2 Animate 14BOffload only Q4_K_MRuns well FP822.6 GB
Wan Animate 2 (14B)Offload only Q4_K_MRuns well Q8_023.4 GB
Wan 2.2 S2V 14BOffload only Q4_K_MRuns well FP821.2 GB
SCAIL-2 (character animation)Offload only Q4_K_MRuns well Q8_023.4 GB
HunyuanVideo (13B, original)Offload only Q4_K_MRuns well Q8_018.3 GB
LTX-Video 13B (0.9.8)Offload only Q2_KRuns well Q8_018.3 GB
HunyuanVideo 1.5Offload only Q4_K_MRuns well 16-bit21.0 GB

“Runs well” means a 16-bit or 8-bit file fits entirely in VRAM.

02A better file, or less offloading

ModelOn the RTX 3070 8 GBOn the RTX 3090 24 GBNeeded
FLUX.2 [dev]Not practical Q2_KRuns Q4_K_M23.4 GB
LTX-2 (19B)Offload only Q4_K_MRuns Q6_K20.8 GB
LTX-2.3 (22B)Offload only Q4_K_MRuns Q6_K22.6 GB
LTX-2.5 (22B)Offload only Q4_K_MRuns Q6_K23.5 GB
MiniMax H3 (33B)Not practical Q3_K_MTight Q3_K_M21.4 GB
MiniMax H3 PrunedOffload only Q4_K_MRuns Q6_K22.5 GB
FLUX.2 [klein] 4BRuns well Q8_0Runs well 16-bit9.6 GB
Mage-Flow (Microsoft)Runs well INT8Runs well 16-bit10.0 GB

03No change in what fits

Same verdict and same file on both cards. Speed can still differ.

04Speed and features

Memory bandwidth goes from 448 to 936 GB/s (×2.09). For models that fit on both cards, bandwidth and compute decide the speed; this ratio is a rough first guide, not a benchmark.