Datasheet for local AI49 models98 GPUsData read 2026-09-25
No adsNo tracking
Check my GPU →

Radeon 8060S (Strix Halo) 96 GB for local AI

Which image and video models run on the Radeon 8060S (Strix Halo) 96 GB, which file to download for each, and how much VRAM they need.

AMD96 GB LPDDR5X (shared system memory) · 256-bit · 256 GB/s · RDNA 3.5Data 2026-09-25
VRAM96 GB
MemoryLPDDR5X (shared system memory)
Bus256-bit
Bandwidth256 GB/s
ArchitectureRDNA 3.5
Launched2025-01
Power—
Runs well49 of 49

Specs: www.amd.com · launch: newsroom.amd.com

01What runs on it

ModelVerdictBest fileSizeNeeded
image models
FLUX.1 [dev]Runs well16-bit23.8 GB26.1 GB
FLUX.1 [schnell]Runs well16-bit23.8 GB26.1 GB
FLUX.1 Kontext [dev]Runs well16-bit23.8 GB26.4 GB
FLUX.1 Krea [dev]Runs well16-bit23.8 GB26.1 GB
FLUX.1 Fill [dev]Runs well16-bit23.8 GB26.4 GB
FLUX.2 [dev]Runs well16-bit64.4 GB67.7 GB
FLUX.2 [klein] 9BRuns well16-bit18.2 GB20.5 GB
FLUX.2 [klein] 4BRuns well16-bit7.8 GB9.6 GB
Krea 2 (Turbo)Runs well16-bit26.3 GB28.9 GB
Qwen-ImageRuns well16-bit40.9 GB43.7 GB
Qwen-Image-Edit (2511)Runs well16-bit40.9 GB43.9 GB
Qwen-Image 2.1Runs well16-bit14.2 GB16.5 GB
Z-Image TurboRuns well16-bit12.3 GB14.3 GB
Z-Image (base)Runs well16-bit12.3 GB14.3 GB
Ideogram 4Runs wellQ8_010.1 GB22.6 GB
Boogu-Image (Turbo)Runs well16-bit20.6 GB23.2 GB
ERNIE-Image (Turbo)Runs well16-bit16.1 GB18.4 GB
HiDream-O1-ImageRuns well16-bit16.4 GB19.2 GB
Mage-Flow (Microsoft)Runs well16-bit8.2 GB10.0 GB
Lumina Image 2.0Runs well16-bit5.2 GB7.0 GB
HiDream-I1 (Full)Runs well16-bit34.2 GB37.0 GB
HiDream-I1 (Dev)Runs well16-bit34.2 GB37.0 GB
Stable Diffusion 3.5 LargeRuns well16-bit16.5 GB18.8 GB
Stable Diffusion 3.5 MediumRuns well16-bit5.1 GB6.9 GB
Chroma1-HDRuns well16-bit17.8 GB20.1 GB
SDXL 1.0Runs well16-bit6.9 GB7.1 GB
Illustrious XL / Pony (SDXL anime)Runs well16-bit6.9 GB7.1 GB
Stable Diffusion 1.5Runs well16-bit2.1 GB3.3 GB
HunyuanImage 2.1Runs well16-bit34.9 GB38.2 GB
video models
Wan 2.1 T2V 14BRuns well16-bit28.6 GB32.9 GB
Wan 2.1 T2V 1.3BRuns well16-bit2.8 GB5.6 GB
Wan 2.1 I2V 14B 480PRuns well16-bit32.8 GB37.1 GB
Wan 2.1 I2V 14B 720PRuns well16-bit32.8 GB39.6 GB
Wan 2.1 VACE 14BRuns well16-bit34.7 GB39.5 GB
Wan 2.2 T2V A14BRuns well16-bit28.6 GB32.9 GB
Wan 2.2 I2V A14BRuns well16-bit28.6 GB32.9 GB
Wan 2.2 TI2V 5BRuns well16-bit10.0 GB13.8 GB
Wan 2.2 Animate 14BRuns well16-bit34.5 GB39.8 GB
Wan Animate 2 (14B)Runs well16-bit32.8 GB38.1 GB
Wan 2.2 S2V 14BRuns well16-bit32.6 GB37.4 GB
SCAIL-2 (character animation)Runs well16-bit32.8 GB38.1 GB
HunyuanVideo (13B, original)Runs well16-bit25.6 GB29.9 GB
LTX-Video 13B (0.9.8)Runs well16-bit28.6 GB32.9 GB
HunyuanVideo 1.5Runs well16-bit16.7 GB21.0 GB
LTX-2 (19B)Runs well16-bit37.8 GB42.6 GB
LTX-2.3 (22B)Runs well16-bit42.0 GB46.8 GB
LTX-2.5 (22B)Runs well16-bit42.0 GB46.8 GB
MiniMax H3 (33B)Runs well16-bit66.3 GB72.1 GB
MiniMax H3 PrunedRuns well16-bit40.2 GB46.0 GB

Calculated from real file sizes plus working memory. How the numbers work.

02Good to know

This is an integrated GPU that uses system memory. AMD states that up to 96 GB of the 128 GB can be set aside as graphics memory on Windows, which is why it can hold models no consumer graphics card can. The catch is speed: the GPU itself (40 compute units) is far smaller than a big desktop card, and its memory bandwidth is a fraction of one, so each step takes much longer. Treat the verdicts here as 'it fits', not 'it is fast'.

03Measured and reported results

LabelModelSetupResultPeak VRAMDateSource
reportedFLUX.1 devflux1-dev full precision (23.8 GB) + t5xxl_fp16 · 1024x1024 · 20 steps
“Steady state | 3.64 s/it | 77.56 s”
AMD ROCm blog, Ryzen AI Max+ 395 / Radeon 8060S, 128 GB unified, Windows ComfyUI; first run 103.89 s; vendor-published measurement
77.56 s / image · 3.64 s/it—2026-07-14rocm.blogs.amd.com →
reportedLTX-2LTX-2 BF16 (T2V) · 1280x720
“"workflow": "LTX2-T2V-BF16.json", ... "duration_seconds": 615.0017409324646”
kyuz0 Strix Halo toolbox benchmark; cold run incl. model load; resolution/frames from benchmark page; I2V = 616.16 s; steps not stated
615 s / clip (121 frames)—2026-02-13raw.githubusercontent.com →
reportedQwen-ImageQwen-Image-2512 BF16 + 4-step Lightning LoRA · 1328x1328 · 4 steps
“"workflow": "Qwen-Image-2512-BF16-4-Step-LoRA.json", ... "duration_seconds": 75.37661480903625”
kyuz0 Strix Halo ComfyUI toolbox benchmark (Ryzen AI Max, ROCm); cold run incl. model load, flags --disable-mmap --gpu-only --disable-smart-memory --cache-none; resolution from ben
75.38 s / image—2026-02-13raw.githubusercontent.com →
reportedQwen-Image-EditQwen-Image-Edit-2511 BF16 + 4-step LoRA · ~1.6MP (dynamic) · 4 steps
“"workflow": "Qwen-Image-Edit-2511-BF16-4-Step-LoRA.json", ... "duration_seconds": 112.70737218856812”
kyuz0 Strix Halo toolbox benchmark; cold run incl. model load; 20-step run = 667.24 s; date from run timestamp
112.71 s / image—2026-02-13raw.githubusercontent.com →
reportedWan 2.2 T2VWan 2.2 T2V A14B (high+low noise experts) · 640x640 · 4 steps
“Steady state | 197 s/it | 186 s/it | 26 min 51 s”
AMD ROCm blog, Ryzen AI Max+ 395; s/it for high-noise / low-noise expert; 26 min 51 s = 1611 s; first run 36 min 14 s; vendor-published
1611 s / clip (81 frames)—2026-07-14rocm.blogs.amd.com →

Reported results are other people's numbers, copied as published, with a link. Settings, drivers and ComfyUI versions differ, so compare them with care. Send yours.