Datasheet for local AI49 models98 GPUsData read 2026-09-25
No adsNo tracking
Check my GPU →

HiDream-O1-Image on 24 GB of VRAM

Image model · 8.8BAny 24 GB GPUData 2026-09-25
Runs well
Yes — and comfortably.

Download 16-bit (16.4 GB). With the model's working memory it needs about 19.2 GB, leaving 4.8 GB spare on 24 GB. Quality: the original weights.

Best file16-bit
File size16.4 GB
VRAM needed~19.2 GB
System RAM32 GB+
Memory map · 24 GB GPU19.2 GB / 24 GB
012 GB24 GB
Weights 16-bit · 16.4 GBWorking memory · 2.0 GBReserve · 0.8 GBFree · 4.8 GB
calculated from real file sizes plus working memory. Not yet measured on this setup. How this works.

01What to download

PartFileFolderSize
Checkpointhidream_o1_image_bf16.safetensors
16-bit · Comfy-Org/HiDream-O1-Image
models/checkpoints16.4 GBDownload →
Total download · keep about the same free on disk16.4 GB

One file: this model has no separate text encoder or VAE. Sizes read from Hugging Face (2026-09-25). “Download” links start the file directly; the file name links are the same files the official ComfyUI workflows use.

System RAM: 32 GB or more. ComfyUI keeps the model file, the text encoder and the VAE in system RAM and moves them to the GPU as needed. With this set of files that is about 16.4 GB, plus roughly 6 GB for Windows, ComfyUI and a browser: 22.4 GB in total. With less RAM it still runs, but Windows starts swapping to disk and loading gets very slow. calculated

02Every file, on 24 GB

FileSizeNeededOn this cardQualityDownload
16-bit ←
SAFETENSORS · Comfy-Org
16.4 GB19.2 GBfits · 4.8 GB sparethe original weightsHugging Face →
FP8
SAFETENSORS · Comfy-Org
8.1 GB10.9 GBfits · 13.1 GB sparepractically identical to the originalHugging Face →

Sizes from the Hugging Face file listing, read 2026-09-25. “Needed” = file + 2 GB working memory for this model + 0.8 GB kept free for the system. The Comfy-Org files are complete checkpoints (loaded from the checkpoints folder); there is no separate text encoder or VAE to add. No GGUF version was found.

03The text encoder

This model has no separate text encoder: reading the prompt is built into the model itself, so its memory is already in the file sizes above.