<?xml version="1.0" encoding="utf-8"?><feed xmlns="http://www.w3.org/2005/Atom"><title>VRAM Ready: new models</title><link href="https://vramready.com/new/"/><link rel="self" href="https://vramready.com/new/feed.xml"/><id>https://vramready.com/new/</id><updated>2026-09-25T00:00:00Z</updated><author><name>VRAM Ready</name></author><entry><title>Qwen-Image 2.1</title><link href="https://vramready.com/models/qwen-image-2-1/"/><id>https://vramready.com/models/qwen-image-2-1/</id><updated>2026-09-01T00:00:00Z</updated><summary>September 2026: a new 7B model that does text-to-image and editing in one, with transparent-background (RGBA) output. Far lighter than the 20B Qwen-Image. Fits from 6 GB of VRAM; 8-bit or better from 10 GB.</summary></entry><entry><title>Wan Animate 2 (14B)</title><link href="https://vramready.com/models/wan-animate-2/"/><id>https://vramready.com/models/wan-animate-2/</id><updated>2026-07-01T00:00:00Z</updated><summary>The July 2026 successor to Wan 2.2 Animate, adding text-driven camera and viewpoint control. The distilled variant runs in about 10 steps. Fits from 12 GB of VRAM; 8-bit or better from 22 GB.</summary></entry><entry><title>MiniMax H3 Pruned</title><link href="https://vramready.com/models/minimax-h3-pruned/"/><id>https://vramready.com/models/minimax-h3-pruned/</id><updated>2026-07-01T00:00:00Z</updated><summary>An official, smaller cut of MiniMax H3 published by Comfy-Org, made for cards that cannot hold the full 33B model. Fits from 15 GB of VRAM; 8-bit or better from 27 GB.</summary></entry><entry><title>MiniMax H3 (33B)</title><link href="https://vramready.com/models/minimax-h3/"/><id>https://vramready.com/models/minimax-h3/</id><updated>2026-07-01T00:00:00Z</updated><summary>The biggest open video release of 2026 and the most downloaded ComfyUI repo of the year: 33B parameters, 4–15 second clips with native stereo sound. Two variants (first/last-frame and reference-to-video); load one. Fits from 22 GB of VRAM; 8-bit or better from 40 GB.</summary></entry><entry><title>Mage-Flow (Microsoft)</title><link href="https://vramready.com/models/mage-flow/"/><id>https://vramready.com/models/mage-flow/</id><updated>2026-07-01T00:00:00Z</updated><summary>Microsoft&#x27;s 4B native-resolution model from July 2026, with separate Turbo (4-step) and Edit checkpoints of the same size. Light enough for most GPUs. Fits from 6 GB of VRAM; 8-bit or better from 6 GB.</summary></entry><entry><title>LTX-2.5 (22B)</title><link href="https://vramready.com/models/ltx-2-5/"/><id>https://vramready.com/models/ltx-2-5/</id><updated>2026-07-01T00:00:00Z</updated><summary>The July 2026 LTX, with a custom Gemma 4 12B text encoder. GGUF sizes differ a lot between uploaders because they mix precisions differently; figures here use realrebelai&#x27;s distilled files. Fits from 14 GB of VRAM; 8-bit or better from 29 GB.</summary></entry><entry><title>SCAIL-2 (character animation)</title><link href="https://vramready.com/models/scail-2/"/><id>https://vramready.com/models/scail-2/</id><updated>2026-06-01T00:00:00Z</updated><summary>June 2026 character animation and replacement built on Wan 2.1 14B: a reference image driven by the motion of another video, without pose skeletons. Native ComfyUI nodes. Fits from 13 GB of VRAM; 8-bit or better from 22 GB.</summary></entry><entry><title>Krea 2 (Turbo)</title><link href="https://vramready.com/models/krea-2/"/><id>https://vramready.com/models/krea-2/</id><updated>2026-06-01T00:00:00Z</updated><summary>Krea&#x27;s own 12B model from June 2026, one of the most downloaded ComfyUI releases of the year. Turbo is the 8-step distilled version; Raw is the base for fine-tuning, same size. Fits from 8 GB of VRAM; 8-bit or better from 16 GB.</summary></entry><entry><title>Ideogram 4</title><link href="https://vramready.com/models/ideogram-4/"/><id>https://vramready.com/models/ideogram-4/</id><updated>2026-06-01T00:00:00Z</updated><summary>Ideogram&#x27;s open-weights model (June 2026, non-commercial). It uses two transformers at once — a main one and an &quot;unconditional&quot; one — so it needs about twice the memory of one file. Fits from 15 GB of VRAM; 8-bit or better from 21 GB.</summary></entry><entry><title>Boogu-Image (Turbo)</title><link href="https://vramready.com/models/boogu-image/"/><id>https://vramready.com/models/boogu-image/</id><updated>2026-06-01T00:00:00Z</updated><summary>A 10B model from June 2026 that generates and edits in one. Turbo needs about 4 steps; Base and the separate Edit checkpoints are the same size. Fits from 11 GB of VRAM; 8-bit or better from 13 GB.</summary></entry><entry><title>HiDream-O1-Image</title><link href="https://vramready.com/models/hidream-o1-image/"/><id>https://vramready.com/models/hidream-o1-image/</id><updated>2026-05-01T00:00:00Z</updated><summary>A 2026 all-in-one 8.8B model built on Qwen3-VL: no separate text encoder and no VAE. It generates, edits and personalises up to 2048×2048. Fits from 11 GB of VRAM; 8-bit or better from 11 GB.</summary></entry><entry><title>ERNIE-Image (Turbo)</title><link href="https://vramready.com/models/ernie-image/"/><id>https://vramready.com/models/ernie-image/</id><updated>2026-04-01T00:00:00Z</updated><summary>Baidu&#x27;s 8B model from April 2026, strong at rendering text in images. Turbo runs in about 8 steps; Base takes 50. Fits from 6 GB of VRAM; 8-bit or better from 11 GB.</summary></entry><entry><title>LTX-2.3 (22B)</title><link href="https://vramready.com/models/ltx-2-3/"/><id>https://vramready.com/models/ltx-2-3/</id><updated>2026-03-01T00:00:00Z</updated><summary>The March 2026 update of LTX-2, grown to 22B parameters. Fits from 14 GB of VRAM; 8-bit or better from 28 GB.</summary></entry><entry><title>Z-Image (base)</title><link href="https://vramready.com/models/z-image/"/><id>https://vramready.com/models/z-image/</id><updated>2026-01-01T00:00:00Z</updated><summary>The undistilled base of Z-Image Turbo: same 6B size, but 28–50 steps with CFG. Better for LoRA training and negative prompts. Fits from 7 GB of VRAM; 8-bit or better from 9 GB.</summary></entry><entry><title>LTX-2 (19B)</title><link href="https://vramready.com/models/ltx-2/"/><id>https://vramready.com/models/ltx-2/</id><updated>2026-01-01T00:00:00Z</updated><summary>Lightricks&#x27; open audio+video model. Generates sound together with the picture. Its text encoder is Gemma 3 12B. Fits from 13 GB of VRAM; 8-bit or better from 26 GB.</summary></entry><entry><title>FLUX.2 [klein] 9B</title><link href="https://vramready.com/models/flux-2-klein-9b/"/><id>https://vramready.com/models/flux-2-klein-9b/</id><updated>2026-01-01T00:00:00Z</updated><summary>A distilled 9B FLUX.2 for fast generation and editing in a few steps. Much lighter than FLUX.2 dev. Fits from 7 GB of VRAM; 8-bit or better from 12 GB.</summary></entry><entry><title>FLUX.2 [klein] 4B</title><link href="https://vramready.com/models/flux-2-klein-4b/"/><id>https://vramready.com/models/flux-2-klein-4b/</id><updated>2026-01-01T00:00:00Z</updated><summary>The smallest FLUX.2: 4B parameters, Apache-2.0 licence, a few steps per image. Fits almost any modern card. Fits from 4 GB of VRAM; 8-bit or better from 6 GB.</summary></entry><entry><title>Qwen-Image-Edit (2511)</title><link href="https://vramready.com/models/qwen-image-edit/"/><id>https://vramready.com/models/qwen-image-edit/</id><updated>2025-12-01T00:00:00Z</updated><summary>The editing version of Qwen-Image, release 2511. Same 20B size and the same text encoder as Qwen-Image. Fits from 11 GB of VRAM; 8-bit or better from 24 GB.</summary></entry><entry><title>Z-Image Turbo</title><link href="https://vramready.com/models/z-image-turbo/"/><id>https://vramready.com/models/z-image-turbo/</id><updated>2025-11-01T00:00:00Z</updated><summary>A 6B, 8-step model from Alibaba&#x27;s Tongyi lab and one of the most downloaded ComfyUI models of the past year. Fast and light. Fits from 6 GB of VRAM; 8-bit or better from 9 GB.</summary></entry><entry><title>HunyuanVideo 1.5</title><link href="https://vramready.com/models/hunyuanvideo-1-5/"/><id>https://vramready.com/models/hunyuanvideo-1-5/</id><updated>2025-11-01T00:00:00Z</updated><summary>Tencent&#x27;s 8.3B video model from late 2025, much lighter than the original 13B HunyuanVideo. The FP8 file is the cfg-distilled variant. Fits from 10 GB of VRAM; 8-bit or better from 13 GB.</summary></entry><entry><title>FLUX.2 [dev]</title><link href="https://vramready.com/models/flux-2-dev/"/><id>https://vramready.com/models/flux-2-dev/</id><updated>2025-11-01T00:00:00Z</updated><summary>The 32B second-generation FLUX. Very capable and very heavy: even 4-bit files are around 20 GB, and the text encoder is a 24B language model. Fits from 17 GB of VRAM; 8-bit or better from 39 GB.</summary></entry><entry><title>Wan 2.2 Animate 14B</title><link href="https://vramready.com/models/wan-2-2-animate-14b/"/><id>https://vramready.com/models/wan-2-2-animate-14b/</id><updated>2025-09-01T00:00:00Z</updated><summary>Character animation and replacement: a reference image of a person is driven by the motion and face of another video. Fits from 12 GB of VRAM; 8-bit or better from 23 GB.</summary></entry><entry><title>HunyuanImage 2.1</title><link href="https://vramready.com/models/hunyuanimage-2-1/"/><id>https://vramready.com/models/hunyuanimage-2-1/</id><updated>2025-09-01T00:00:00Z</updated><summary>Tencent&#x27;s 17B model that generates natively at 2K. Its GGUF files come from calcuis and need calcuis&#x27;s own GGUF loader node. Fits from 11 GB of VRAM; 8-bit or better from 21 GB.</summary></entry><entry><title>Wan 2.2 S2V 14B</title><link href="https://vramready.com/models/wan-2-2-s2v-14b/"/><id>https://vramready.com/models/wan-2-2-s2v-14b/</id><updated>2025-08-01T00:00:00Z</updated><summary>Sound-to-video: a talking or singing person from one image plus an audio track. Fits from 15 GB of VRAM; 8-bit or better from 22 GB.</summary></entry><entry><title>Qwen-Image</title><link href="https://vramready.com/models/qwen-image/"/><id>https://vramready.com/models/qwen-image/</id><updated>2025-08-01T00:00:00Z</updated><summary>Alibaba&#x27;s 20B text-to-image model, known for rendering long text and posters well. Big: the 8-bit file alone is over 20 GB. Fits from 10 GB of VRAM; 8-bit or better from 24 GB.</summary></entry><entry><title>Chroma1-HD</title><link href="https://vramready.com/models/chroma1-hd/"/><id>https://vramready.com/models/chroma1-hd/</id><updated>2025-08-01T00:00:00Z</updated><summary>An 8.9B community model built from FLUX.1 schnell, Apache-2.0 licensed, with a big fine-tuning scene. Fits from 6 GB of VRAM; 8-bit or better from 12 GB.</summary></entry><entry><title>Wan 2.2 TI2V 5B</title><link href="https://vramready.com/models/wan-2-2-5b/"/><id>https://vramready.com/models/wan-2-2-5b/</id><updated>2025-07-01T00:00:00Z</updated><summary>A single dense 5B model for text- or image-to-video at 720p. The friendliest Wan for 8–12 GB cards. Fits from 6 GB of VRAM; 8-bit or better from 10 GB.</summary></entry><entry><title>Wan 2.2 T2V A14B</title><link href="https://vramready.com/models/wan-2-2-t2v-14b/"/><id>https://vramready.com/models/wan-2-2-t2v-14b/</id><updated>2025-07-01T00:00:00Z</updated><summary>Wan 2.2 splits the work between two 14B models: one for the early, noisy steps and one for the late ones. Only one needs to be in VRAM at a time, but both must fit in system RAM. Fits from 10 GB of VRAM; 8-bit or better from 19 GB.</summary></entry><entry><title>Wan 2.2 I2V A14B</title><link href="https://vramready.com/models/wan-2-2-i2v-14b/"/><id>https://vramready.com/models/wan-2-2-i2v-14b/</id><updated>2025-07-01T00:00:00Z</updated><summary>The image-to-video Wan 2.2, and the most downloaded Wan 2.2 GGUF. Same two-model setup as the text-to-video version. Fits from 10 GB of VRAM; 8-bit or better from 19 GB.</summary></entry><entry><title>LTX-Video 13B (0.9.8)</title><link href="https://vramready.com/models/ltx-video-13b/"/><id>https://vramready.com/models/ltx-video-13b/</id><updated>2025-07-01T00:00:00Z</updated><summary>The last of the pre-LTX-2 video models from Lightricks (July 2025). No audio, T5 text encoder, and much lighter than LTX-2. Fits from 10 GB of VRAM; 8-bit or better from 19 GB.</summary></entry></feed>