Datasheet for local AI49 models98 GPUsData read 2026-09-25
No adsNo tracking
Check my GPU →

RTX 4060 Ti 16 GB for local AI

Which image and video models run on the RTX 4060 Ti 16 GB, which file to download for each, and how much VRAM they need.

NVIDIA16 GB GDDR6 · 128-bit · 288 GB/s · Ada LovelaceData 2026-09-25
VRAM16 GB
MemoryGDDR6
Bus128-bit
Bandwidth288 GB/s
ArchitectureAda Lovelace
Launched2023-07
Launch price$499
Runs well25 of 49

Specs: www.nvidia.com · launch: en.wikipedia.org

01What runs on it

ModelVerdictBest fileSizeNeeded
image models
FLUX.1 [dev]Runs wellFP811.9 GB14.2 GB
FLUX.1 [schnell]Runs wellFP811.9 GB14.2 GB
FLUX.1 Kontext [dev]Runs wellFP811.9 GB14.5 GB
FLUX.1 Krea [dev]Runs wellFP811.9 GB14.2 GB
FLUX.1 Fill [dev]Runs wellQ8_012.7 GB15.3 GB
FLUX.2 [dev]Offload onlyQ2_K12.9 GB16.2 GB
FLUX.2 [klein] 9BRuns wellFP89.4 GB11.7 GB
FLUX.2 [klein] 4BRuns well16-bit7.8 GB9.6 GB
Krea 2 (Turbo)Runs wellFP813.1 GB15.7 GB
Qwen-ImageRunsQ4_K_M13.1 GB15.9 GB
Qwen-Image-Edit (2511)TightQ3_K_M9.9 GB12.9 GB
Qwen-Image 2.1Runs wellQ8_07.6 GB9.9 GB
Z-Image TurboRuns well16-bit12.3 GB14.3 GB
Z-Image (base)Runs well16-bit12.3 GB14.3 GB
Ideogram 4RunsQ4_16.2 GB14.7 GB
Boogu-Image (Turbo)Runs wellFP810.3 GB12.9 GB
ERNIE-Image (Turbo)Runs wellQ8_08.7 GB11.0 GB
HiDream-O1-ImageRuns wellFP88.1 GB10.9 GB
Mage-Flow (Microsoft)Runs well16-bit8.2 GB10.0 GB
Lumina Image 2.0Runs well16-bit5.2 GB7.0 GB
HiDream-I1 (Full)RunsQ5_K_M13.0 GB15.8 GB
HiDream-I1 (Dev)RunsQ5_K_M13.0 GB15.8 GB
Stable Diffusion 3.5 LargeRuns wellQ8_08.8 GB11.1 GB
Stable Diffusion 3.5 MediumRuns well16-bit5.1 GB6.9 GB
Chroma1-HDRuns wellFP89.2 GB11.5 GB
SDXL 1.0Runs well16-bit6.9 GB7.1 GB
Illustrious XL / Pony (SDXL anime)Runs well16-bit6.9 GB7.1 GB
Stable Diffusion 1.5Runs well16-bit2.1 GB3.3 GB
HunyuanImage 2.1RunsQ4_K_M11.3 GB14.6 GB
video models
Wan 2.1 T2V 14BRunsQ5_K_M11.3 GB15.6 GB
Wan 2.1 T2V 1.3BRuns well16-bit2.8 GB5.6 GB
Wan 2.1 I2V 14B 480PRunsQ4_K_M11.3 GB15.6 GB
Wan 2.1 I2V 14B 720PTightQ3_K_M8.6 GB15.4 GB
Wan 2.1 VACE 14BTightQ3_K_S7.8 GB12.6 GB
Wan 2.2 T2V A14BRunsQ5_K_M10.8 GB15.1 GB
Wan 2.2 I2V A14BRunsQ5_K_M10.8 GB15.1 GB
Wan 2.2 TI2V 5BRuns well16-bit10.0 GB13.8 GB
Wan 2.2 Animate 14BTightQ3_K_M8.6 GB13.9 GB
Wan Animate 2 (14B)TightQ3_K_M8.6 GB13.9 GB
Wan 2.2 S2V 14BTightQ2_K9.5 GB14.3 GB
SCAIL-2 (character animation)TightQ3_K_M9.1 GB14.4 GB
HunyuanVideo (13B, original)RunsQ6_K11.0 GB15.3 GB
LTX-Video 13B (0.9.8)RunsQ6_K10.9 GB15.2 GB
HunyuanVideo 1.5Runs wellFP88.3 GB12.6 GB
LTX-2 (19B)TightQ3_K_M10.1 GB14.9 GB
LTX-2.3 (22B)TightQ3_K_M10.8 GB15.6 GB
LTX-2.5 (22B)TightQ2_K8.8 GB13.6 GB
MiniMax H3 (33B)Offload onlyQ4_K_M19.9 GB25.7 GB
MiniMax H3 PrunedTightQ3_K_M8.9 GB14.7 GB

Calculated from real file sizes plus working memory. How the numbers work.

02Good to know

The RTX 4060 Ti 16 GB is an Ada Lovelace GPU with hardware FP8, so ComfyUI can compute Comfy-Org's FP8 files natively: small and fast. (Plain FP8 files use FP8 maths with the --fast fp8_matrix_mult option.)

03Measured and reported results

LabelModelSetupResultPeak VRAMDateSource
reportedFLUX.1 devFP8
“Flux-Dev(FP8) … 2回目は72秒から51秒とあまり効果が感じられませんでした。”
ComfyUI; same PC before/after swapping RTX 4060 Ti 8GB -> 16GB; 2nd-run times; resolution/steps not stated; 51 s = 16GB card
51 s / image—2025-01-03note.com →
reportedFLUX.1 schnellFP8
“Flux-Schnell(FP8) … 2回目は24秒から11秒です。”
ComfyUI; same PC before/after swapping RTX 4060 Ti 8GB -> 16GB; 2nd-run times; resolution/steps not stated; 11 s = 16GB card
11 s / image—2025-01-03note.com →
reportedQwen-Image 2.1Qwen Image 2.1 INT8 Convrot (+Qwen3 VL 8B INT8 TE)
“Image generation: ~20 seconds ... Peak VRAM: ~15 GB”
ComfyUI, no CPU/disk offload; approximate values ('~'); editing ~60 s; resolution/steps not stated; date derived from '1 day ago' on 2026-09-25
20 s / image15 GB2026-09-24huggingface.co →

Reported results are other people's numbers, copied as published, with a link. Settings, drivers and ComfyUI versions differ, so compare them with care. Send yours.