Previous Post

SwarmUI Auto Installer + The Ultimate Image and Video AI Models Downloader - For Windows, RunPod and Massed Compute - Ultimate Compilation

Next Post
SwarmUI Auto Installer + The Ultimate Image and Video AI Models Downloader - For Windows, RunPod and Massed Compute - Ultimate Compilation
1 / 31
DESCRIPTION

Automatic SwarmUI Installer with Extra Extensions + Unified AI Models Downloader for SwarmUI, ComfyUI, Automatic1111 and Forge Web UI, Forge Web UI Classic and Neo - Supports SD 1.5, SDXL, FLUX, FLUX Krea, FLUX SRPO, Wan 2.1, Wan 2.2, Qwen Image, Qwen Image Edit Plus 2509, LTX 2.3, Krea 2, Full Int8 ConvRot Quants + SwarmUI presets to generate images and videos and more

Patreon exclusive posts index to find our scripts easily, Patreon scripts updates history to see which updates arrived to which scripts and amazing Patreon special generative scripts list that you can use in any of your task.

Join discord to get help, chat, discuss and also tell me your discord username to get your special rank : SECourses Discord

Please also Star, Watch and Fork our Stable Diffusion & Generative AI  GitHub repository and join our Reddit subreddit and follow me on LinkedIn (my real profile)

=======

Latest zip file : SwarmUI_Model_Downloader_v154.zip

[click here to choose a membership and Join to download zip files]

Main Install Tutorial : https://youtu.be/c3gEoAyL2IE

Install Tutorial on Cloud (RunPod + MassedCompute) : https://youtu.be/bBxgtVD3ek4

SimplePod Tutorial : https://youtu.be/yOj9PYq3XYM

SimplePod Register : https://simplepod.ai/ref?user=secourses

Qwen Image Models Training Tutorial : https://youtu.be/DPX3eBTuO_Y

ComfyUI Back-end Installer with Torch 2.13 CUDA 13 and xFormers, TorchAO, Sage Attention, Segment Anything, Segment Anything 2 (SAM2), MSLK (xFormers won't work properly without this), InsightFace, Flash Attention https://www.patreon.com/posts/105023709

Definitely use our ComfyUI solo backend installer otherwise some of the newest stuff may fail

Example existing ComfyUI backend : E:\Comfy_UI_V97\ComfyUI\main.py

Sage Attention is optional

If you use Sage Attention and get black output, enable Display Advanced Options of SwarmUI and go Advanced Sampling and make Preferred DType = Default (16 bit)

Highest quality attention is --use-pytorch-cross-attention

New recommended backend is --enable-triton-backend

Our Sage Attention is newest compiled and should not produce low quality or black outputs

If your GPU starts using shared VRAM for any reason, add this command to make it avoid that --reserve-vram 3

This command will preserve 3 GB VRAM for other tasks

--cache-none is useful when you work with dual models or keep switching between models

--novram is useful when you are getting OOM, especially longer video generation like LTX 2, it reduces VRAM usage and can be used together with --cache-none

Some VRAM savings --disable-smart-memory, --lowvram

Use --disable-smart-memory if you are getting stuck or OOM

Image and Video slider app : https://www.patreon.com/posts/133935178

19 July 2026 V154

This is a very big update with so many important new features, changes and improvements so please read all thank you

Seven New Presets

Make sure to import and overwrite Amazing_SwarmUI_Presets_v59.json or use Windows_Preset_Delete_Import.bat while SwarmUI running

Krea 2

Krea 2 Raw/Base Slow - 260711 uses 52 steps, Euler, Simple, and CFG 4.5. It is the slower Raw/Base path intended mainly for fine-tuning, post-training, and LoRA work.

Krea 2 Turbo Fast - 260711 uses 8 steps, Euler, Beta, and CFG 1 for fast final inference.

Krea 2 Turbo Image Edit - 260716 adds denoise-based image editing with a default creativity value of 0.65. Add the source as Init Image and describe the requested change.

The new Krea 2 Core Bundle contains the Raw/Base INT8 HQ model, Turbo INT8 HQ model, Qwen3-VL 4B text encoder, and Qwen Image VAE. Its cached total is approximately 34.83 GiB.

The catalog also adds individually selectable Krea 2 Turbo Q8 GGUF and NVFP4 High Quality variants.

Automatically installed ComfyUI-QuantOps updated (With our ComfyUI Installer for SwarmUI Backend)

Now it is only used if native ComfyUI is not supporting the loaded model

I have compared the new famous Int8 ConvRot of Krea 2 and the speed difference is like 100%, I plan to update all models to Int8 ConvRot

New updated ComfyUI-QuantOps supports Krea 2 GGUF as well

When you click and see full size of above image and analyze results you will see that:

Int8 ConvRot is 96.2% similar to BF16 meanwhile GGUF Q8 is only 90.0% and FP8 Scaled is 82.2% and NVFP4 is 63.7%

Moreover, Int8 ConvRot generates the output in 3.05 seconds, making it 1.82× faster than BF16, which takes 5.56 seconds.

NVFP4 takes 3.8 seconds and is 1.46× faster than BF16, whereas GGUF Q8 takes 6.06 seconds and is approximately 8.3% slower than BF16.

So Int8 ConvRot generated with our Musubi Trainer app at high quality is almost 100% faster and almost same quality as BF16

High quality generation takes few hours on RTX 5090

Updated SwarmUI downloader has amazing quality Int8 ConvRot which reaches almost BF16 quality : https://www.patreon.com/SECourses/posts/swarmui-auto-and-114517862

Krea 2 Core bundle downloads this model and uses it in SwarmUI preset

ComfyUI preset also uses that model default

You can generate Int8 ConvRot models with our updated Musubi Trainer app : https://www.patreon.com/SECourses/posts/secourses-musubi-137551634

LTX 2.3 Foley Video-to-Audio

The new LTX23 Foley Video To Audio preset generates synchronized audio for a silent or muted video while keeping the source video unchanged.

Put the video in Init Image, then describe the visible sound source, action, material, and timing. The preset uses the LTX 2.3 Foley V2A LoRA with 30 steps, audio CFG 6, STG scale 1, modality scale 3, and a maximum of 169 frames at 24 FPS.

The Foley LoRA is approximately 0.21 GiB and is now part of the LTX 2.3 Video Core Bundle. The updated premium installer also installs the maintained FoleyExtension node package automatically.

This model uses Dev version of LTX 2.3 not Turbo, therefore, now Dev version is included in the LTX 2.3 core bundle

For clean Foley generation, a useful prompt ending is: No speech is present. No music is present.

Download LTX 2.3 Core Bundle and in ComfyUI, install bundle 100

LTX 2.3 Licon MSR V2

The new LTX23 Licon MSR V2 Multi Subject Reference preset supports multiple subject, object, texture, or viewpoint references plus a required background image.

Add two to five images to Prompt Images in this exact order:

One to four subject, object, texture, or viewpoint references (so first 1-4 images are subjects)

The required background image last (last image is always background)

Therefore, you need minimum 2 input images into the prompt field

Do not use Init Image for this workflow. Identify each numbered reference and its role in the positive prompt, then describe the new action, scene, camera, and lighting.

The preset uses the official eight-step distilled sigma schedule, Euler Ancestral sampling, a 65-frame internal reference sequence, and the LTX 2.3 Licon MSR V2 IC-LoRA. The new LoRA is approximately 0.61 GiB and is also included in the LTX 2.3 Video Core Bundle.

A dedicated guide is included with reference ordering, recommended settings, installation paths, and licensing details.

For this to work, we have coded custom extensions and all is automatically installer with SwarmUI installer or Updater files

Download LTX 2.3 Core Bundle and in ComfyUI, install Bundle 100 for ComfyUI backend

It is insanely powerful and fast I mean look below example

LTX 2.3 Presets Updates

LTX 2.0 model presets removed since LTX 2.3 is better in everyway

LTX 2.3 presets now using myself compiled Int8 ConvRot HQ models since they are literally 100% faster

Model downloader LTX 2.3 core bundle is now downloading these new models

You can generate Int8 ConvRot models with our updated Musubi Trainer app : https://www.patreon.com/SECourses/posts/secourses-musubi-137551634

So Int8 ConvRot HQ is 100% faster than FP8 Quant Scaled and 50% faster than BF16 on RTX 5090

The quality is also excellent almost same as BF16

Phantom-Wan Character Reference T2V

Two Phantom-Wan presets are included:

Phantom Wan 14B Character Reference T2V Fast

Phantom Wan 14B Character Reference T2V Quality

Add character references to Prompt Images, not Init Image, then describe every subject, the scene, action, camera, and lighting. The reference images condition identity while generation begins from an empty video latent, so this is true reference-conditioned text-to-video rather than first-frame image-to-video.

One to four references are the officially recommended path. The integration can accept up to six for experimentation.

Download Phantom Wan 14B Character Reference T2V Bundle and in ComfyUI, install Bundle 100 for ComfyUI backend

Bundle 100 for ComfyUI Backend

Inside ComfyUI zip file : Windows_Custom_Nodes_Bundles_Installer.bat

Unified Fast Robust Model Downloader Improvements

Run Windows_Start_Download_Models_App.bat to start Downloader

The previous hf_transfer checkbox has been replaced with two clearer controls:

Custom Downloader Threads: 1 to 32, with 16 as the recommended default

Hugging Face Xet Backend: optional and disabled by default

The custom thread setting applies to individual catalog downloads, bulk queues, bundles, search results, snapshots, and direct URL downloads. When Xet is enabled, the custom thread setting is ignored. If Xet fails, v150 falls back to the custom resumable backend unless the operation was cancelled.

The custom backend now has much stronger recovery behavior:

Known-size partial downloads are preserved when cancelled.

Parallel downloads save their range layout so later attempts can resume safely.

Legacy v145 16-part downloads can be migrated and resumed.

Every retry re-reads the saved byte count to prevent duplicate ranges and corruption.

Six attempts use bounded exponential backoff.

HTTP 429, 500, 502, 503, and 504 responses receive controlled retries.

Separate connection and idle-read timeouts recover stalled transfers sooner.

Progress output reports when it is waiting for the server and uses smoother speed estimates.

Free space is checked before downloading and before merging parallel parts, with a safety reserve.

Disk-full failures preserve useful partial data.

Parallel parts are merged to a temporary file and atomically activated.

Hugging Face main file URLs can use SHA verification and the optional Xet path.

SHA-256 verification and the verified-file cache remain in place. The goal is simple: fewer restarts from zero, clearer logs, and safer recovery on large model downloads.

Now you can set Hugging Face token for faster or private repo downloads

Now you can set HF Xet Download - even faster optional

Now you can set number of download threads

Now you can set Parallel File Downloads count

This is extremely useful on high bandwith systems to download even faster than Hugging Face single file download limits like reaching 1 GB per second

Installation and Update Reliability

Windows install and update scripts no longer require one exact .NET SDK patch version. The new helper searches common system, user, registry, PATH, and environment locations for any usable stable .NET 10.x SDK while ignoring preview-only builds.

If installation is needed, it tries WinGet first and then falls back to the latest official Microsoft installer for x64, ARM64, or x86. The fallback download is checked against Microsoft's published release hash before it runs, and installer progress and log locations are shown instead of failing silently.

The premium-extension installer now validates the required integrations before changing the active setup. It installs or updates:

FoleyExtension nodes plus the managed LTX 2.3 Foley SwarmUI integration

Phantom character-reference helper nodes for the Comfy backend

The managed Licon MSR SwarmUI integration

The official Licon MSR Comfy node at the tested pinned revision

The new helper-node installers use staged directories, required-file checks, backup/rollback paths, bounded Git operations, and stale compiled-extension cleanup. Windows, RunPod, and Massed Compute update instructions now stop with an explicit error if this stage fails instead of continuing into a partial setup.

Updated Preset Model Report

A new file, Windows_Update_And_Open_Model_Report.bat, regenerates and opens the HTML report that shows which bundles and model files cover every SwarmUI preset.

The report generator now handles Krea 2, Phantom-Wan, Ideogram 4, negative-model parameters, saved-filename aliases, and LTX 2.3 Dev versus Distilled connector selection more accurately. The packaged report has been refreshed against Amazing_SwarmUI_Presets_v55.json.

How to Upgrade

Existing installation

Close SwarmUI and the model downloader.

Back up custom presets and any source changes you made inside the SwarmUI Git checkout. The updater resets tracked checkout files before pulling upstream changes.

Extract the latest zip file over your existing downloader package folder and allow release files to be replaced. Do not delete your SwarmUI or model directories.

Run Windows_Update_SwarmUI.bat.

Start the model downloader and download only the new bundles or individual files you need.

Import Amazing_SwarmUI_Presets_v55.json.

Windows_Preset_Delete_Import.bat can automate the preset replacement while SwarmUI is running. It first backs up the current presets, then deletes all active presets and imports the latest pack. Read its confirmation carefully if you maintain custom presets.

Users who already have the rest of LTX 2.3 can download only the new Foley and Licon LoRAs instead of downloading the full LTX bundle again.

Run Windows_Update_And_Open_Model_Report.bat afterward to compare the new preset requirements with the models currently on disk.

Fresh installation

Extract latest zip file into a clean folder.

Run Windows_Install_SwarmUI.bat.

Start the model downloader with Windows_Start_Download_Models_App.bat.

Download the bundle or individual models needed by your chosen presets.

Start SwarmUI and import latest Amazing_SwarmUI_Presets_.json

RunPod and Massed Compute users should follow the updated instruction files included in the package.

2 July 2026 V145

SwarmUI going to require .NET SDK 10 therefore, both installer and bat file will auto install .NET SDK 10 if you are missing

It will ask permission to install so click yes / ok

RunPod, SimplePod, Massed Compute installers also updated to auto install, no need confirm there

Ideogram 4 models added into Image Generation Models tab in model downloader app

For JSON prompt generation for Ideogram 4 model, we have ultimate new app : https://www.patreon.com/SECourses/posts/162527725

Ideogram 4 Core Bundle added into SwarmUI Bundles tab in model downloader app

With Amazing_SwarmUI_Presets_v51.json 3 new presets added as below

Ideogram 4 Highest Quality - 260629.json

Ideogram 4 Balanced - 260629.json

Ideogram 4 Turbo - 260629.json

All 3 preset uses Ideogram 4 unconditional model as well for high quality but you can disable if you want

I think when you disable it becomes more non-realistic

Ideogram 4 Turbo generates amazing images at 2048x2048 and really fast speed

Use Windows_Preset_Delete_Import.bat to fresh update

If you have custom presets they will be deleted be careful

You can also import Amazing_SwarmUI_Presets_v51.json and just import new ones

Updated new all presets

26 April 2026 V142

New LoRA LTX2.3_Crisp_Enhance.safetensors (0.66 GB) added to the LTX 2.3 Core bundle in our model downloader app

All SwarmUI LTX 2.3 presets updated to use this new LoRA which improves quality significantly

Get latest zip file and get v50 preset file

Which_Bundles_Downloads_Which_Preset_Models.html updated

23 April 2026 V141

LTX 2.3 Video Core Bundle updated and the following new LoRAs added

LTX23_Anime-To-Real_v1.safetensors (0.61 GB)

LTX23_Edit-Anything_v1.safetensors (1.22 GB)

LTX23_Real-To-Anime_v1.safetensors (0.61 GB)

Default model replaced to LTX 2.3 v1.1 : LTX2.3-Distilled-v1.1-FP8-Quant-Scaled.safetensors

I have converted it to Quant Scaled very high quality

LTX 2.3 presets updated to use v1.1 model as default

New preset LTX23 Video To Video 8 Steps added

This preset takes video as input instead of image from same place

It automatically recognize it is video or image don't worry

By using above LoRAs or even without LoRA, you can change existing videos

Play prompt and Init Image Creativity level (default 0.5) to get your desired results

Example simple prompt : Convert the video into a anime style

New updated presets are provided with Amazing_SwarmUI_Presets_v49.json

I recommend to use Windows_Preset_Delete_Import.bat file

Our model downloader latest statistics are as below

22 April 2026 V140

New precision filters added

New FLUX 2 Klein Models (11 models - Total: 105.17 GB) and ERNIE Image Models (7 models - Total: 59.45 GB) added into Image Generation Models

New bundles added

FLUX 2 Klein Core Bundle (Total: 34.74 GB, 5 models)

ERNIE Image Core Bundle (Total: 23.12 GB, 4 models)

I have compiled myself the following models for you via our SECourses Musubi Tuner model quantizer so you get them first time in the entire community - https://www.patreon.com/posts/137551634

ernie-image-turbo-Quant-FP8-Scaled.safetensors (7.81 GB)

ernie-image-Quant-FP8-Scaled.safetensors (7.81 GB)

FLUX-2-Klein-Base-9b-Quant-FP8.safetensors (8.79 GB)

FLUX-2-Klein-Distilled-Quant-FP8-Scaled-9b-kv.safetensors (8.79 GB)

FLUX-2-Klein-Distilled-9b-Quant-FP8-Scaled.safetensors (8.79 GB)

Quant FP8 Models are all tested and excellent quality

After extensive testing on 4x RTX PRO 6000 GPUs having Massed Compute machine, new following presets added to the our arsenal with Amazing_SwarmUI_Presets_v48.json

ERNIE Image Base - 260422

ERNIE Image Turbo - 260422

FLUX 2 Klein Base - 260422

FLUX 2 Klein Distilled 8 Steps - 260422

You can use Windows_Preset_Delete_Import.bat or manually import new presets

For presets to work, download FLUX 2 Klein Core Bundle (Total: 34.74 GB, 5 models) and ERNIE Image Core Bundle (Total: 23.12 GB, 4 models)

Tutorials tab on the downloader app updated

Get latest zip file and overwrite older files to have updated files

30 March 2026 V138

I have compiled and added LTX2.3-Dev-FP8-Quant-Scaled.safetensors (27.14 GB) to model list

SwarmUI presets updated according to the latest SwarmUI version so you won't get warnings, i recommend overwrite all others via Windows_Preset_Delete_Import.bat

Gradio version of the app upgraded now even better and faster

16 March 2026 V137

LTX 2.3 models bundle updated for ComfyUI presets - use ComfyUI latest zip file to get them

LTX2.3_Enchance_Prompt_Feed_For_LLMs.txt significantly improved

Works great for Text to Video and Image to Video prompt generation

Upload it to your favorite LLM e.g. ChatGPT, provide input image or whatever you want and describe and tell LLM to generate / improve prompt

13 March 2026 V135

Model downloader app upgraded to Gradio 6.9.0 and working blazing fast now

New LTX 2.3 Video Models (26 models - Total: 275.22 GB) added

Some model sizes were wrongly displayed and it is fixed

Which_Bundles_Downloads_Which_Preset_Models.html file updated

New LTX 2.3 Video Core Bundle (Total: 67.70 GB, 15 models) added please download this

I have myself compiled the very best FP8 Quant Scaled LTX 2.3 model

At the moment GGUF of LTX 2.3 not working in SwarmUI at the moment since ComfyUI didn't add native support yet, thus i recommend using FP8 Quant Scaled should work amazing if you have RAM memory

We still have GGUF models in downloader app and I plan to update ComfyUI with presets for GGUF

Our ComfyUI already supports GGUF of LTX 2.3

New 2 presets added with Amazing_SwarmUI_Presets_v46.json

LTX23 Image To Video 8 Steps - 260310

LTX23 Text To Video 8 Steps - 260310

No LoRAs used at the moment but I may update in future if there be nice LoRAs

Atm working amazing even without any LoRAs

Download latest zip file and overwrite all files in install folder and use newest model downloader app and import new presets

28 February 2026 V133

Gradio version upgraded to 6.8.0 now downloader app runs faster

RIFE frame interpolation repo was outdated, i forked and fixed it in SwarmUI installers

There was a bug when downloading models from CivitAI and now it is made more robust and error fixed

14 February 2026 V131

This update was for our Ultimate Unified AI Models downloader app

Massive revamp and improvement made to the User Interface (UI) and User Experience (UX)

We have at least 10x improved the app performance and responsiveness of the UI

I also have reported a bug to Gradio team to even further improve performance hopefully

https://github.com/gradio-app/gradio/issues/12891

URL downloader significantly improved and made more robust

Tutorials section fully updated

Check the below screenshots to see and please try every feature and tab and let me know what else you want

For updating as usual download the zip file and overwrite all of the previous files in your installation folder

3 February 2026 V130

Model downloader app interface significantly modernized and improved

CMD download logs and messages improved

LTX2_Enchance_Prompt_Feed_For_LLMs.txt added into zip file which you can upload to https://aistudio.google.com/prompts/new_chat and write your prompt and that is it. You can also upload image + this file and prompt to enhance your prompt for free

Following new presets added into our amazing SwarmUI presets with Amazing_SwarmUI_Presets_v45.json - so update your presets

LTX2 Image To Video 8 Steps

Image to video works great and upscale normally not supported

For this preset to make also LTX 2 upscale work, I have developed our first SwarmUI premium extension and now it is auto installed with SwarmUI installers and also update bat files

Also for upscale, you need to set your base resolution half and apply upscale like 480x480 and 2x upscaled to 960x960

I am using distilled FP8 Scaled model directly thus it is 8 steps

LTX2 Text To Video 8 Steps

Working amazing and you can use it with LTX2 Video Upscale preset

To use with LTX2 Video Upscale preset, reduce your base resolution to half like 480x480 instead of 960x960

Z Image Base preset added and working amazing new Z Image base model

The preset is 40 steps but you can do lesser if you want

Also i recommend you to combine it with Upscale Images 2X preset to get very high res and quality images

Currently Z Image doen't work accurately with --use-sage-attention so you can replace attention with --use-pytorch-cross-attention

For the model downloader app the following changes made

Z Image Turbo Core Bundle renamed into Z Image Core Bundle

For Z Image presets you need this bundle

Z Image BF16 model added into Z Image Core Bundle - this is new today published base model and works great

Z Image Turbo Models renamed into Z-Image Models

Z Image BF16 (11.46 GB) and Z Image Quant FP8 Scaled (5.86 GB) models added

I have compiled myself the Z Image Quant FP8 Scaled and its quality almost same

I have used our SECourses Musubi Trainer app with highest quality Quant FP8 Scaled Maker

LTX 2 Video Models added to the downloader app

LTX 2 Video Core Bundle added to the app

For LTX 2 presets you need this bundle

If you get OOM with LTX 2 presets, add the following arguments into your ComfyUI backend and test

--lowvram or --novram

Also you can try --disable-smart-memory

All these can be also combined with --cache-none : lowest RAM

LTX 2 Video Models also has LTX-2-19b-Dev-GGUF_Q4_K_M (12.46 GB) and gemma_3_12B_it_fp4_mixed.safetensors (8.80 GB) for low VRAM + low RAM machines

Complete Image Generation and Editing Bundle now includes new Z Image BF16 model as well

I am working on a tutorial video to explain all hopefully

And hopefully I will update our ComfyUI installer as well to add there too

21 January 2026 V127

MassedCompute_SwarmUI_Install_Instructions-Optional.txt added

This is optional since we have SwarmUI already in Massed Compute as shown in tutorials

You can use this to install SwarmUI in Linux machines

I tested and it works perfect

15 January 2026 V126

I have updated the NVFP4 quantizer app and implemented Mixed NVFP4 for FLUX models

With this change, new FLUX SRPO Mixed NVFP4 model added into app which has amazing quality : https://imgsli.com/NDQyNjk5

Mixed NVFP4 speed is same and size is slightly bigger than full NVFP4

14 January 2026 V125

I have developed a NVFP4 quantizer app and published here : https://www.patreon.com/posts/148217625

With using it I have generated FLUX SRPO NVFP4 and it is amazing for low VRAM GPUs and ultra fast

FLUX SRPO NVFP4 added into the downloader app into FLUX Models section

We are the first one to compile and publish it

Hopefully I will work on Qwen 2512 later

13 January 2026 V124

New model FLUX 1 Dev Kontext Quant FP8 Scaled added to the downloader app

I compiled it myself and amazing quality

ComfyUI is now supporting NVFP4 LoRAs not perfect and not all models but i believe it will hopefully get better, I tested on FLUX 2, FLUX 1 and Z Image Turbo

FLUX 2 failed, FLUX 1 i think worked nice, Z Image Turbo quality dropped

RunPod template link updated and now we fully support SimplePod which is much faster and cheaper than RunPod

SIMPLEPOD CHEAPER AND FASTER THAN RUNPOD

Now we fully support SimplePod as well please use this link to register : https://simplepod.ai/ref?user=secourses

SimplePod is faster and cheaper than RunPod and works exactly same

E.g. RTX 5090 on RunPod is 0.89 USD per hour, on SimplePod it is 0.45$ per hour,

RTX PRO 6000 on RunPod is 1.84 USD per hour and on SimplePod it is 0.79 USD per hour

Please use this template on SimplePod : https://dash.simplepod.ai/account/explore/100/ref-secourses/

For permanent storage, generate it from Storage tab with any name and size you want and when selecting template with above link, click Edit and Use, select Persistence Volume and change mount point to /workspace

Up-to-date SimplePod tutorial starting from 21:51 : https://youtu.be/yOj9PYq3XYM?si=Z86wZZLBeYzWo1Qo&t=1311

As usual follow Massed_Compute_Instructions_READ.txt and RunPod_SimplePod_Instructions_READ.txt to install and use and watch the tutorials

10 January 2026 V123

Full tutorial published : https://youtu.be/yOj9PYq3XYM

We have added NVFP4 models both into their respected Model category and as a bundle

FLUX 2 Dev NVFP4 (21.21 GB)

FLUX Dev Kontext NVFP4 (8.56 GB)

FLUX Dev NVFP4 (8.56 GB)

Z Image Turbo NVFP4 (4.20 GB)

NVFP4 models uses same presets as FP8 Scaled or BF16

Sadly, currently, LoRAs not working with NVFP4

I always recommend FP8 Scaled over GGUF models if your RAM is sufficient to do Block Swapping / VRAM Streaming

Use --disable-smart-memory args in your ComfyUI backend if you are getting stuck or OOM - this reduces VRAM usage but a little bit slower

To use NVFP4 models, you have to install latest V73 or above ComfyUI read its announcement : https://www.patreon.com/posts/105023709

I have compiled Quant FP8 Scaled of FLUX 1 Dev and added to downloads

Moreover, for lower VRAM machines, Z Image Turbo now has the following text encoder models

Qwen 3 4B Text Encoder FP8 Mixed (For Z Image Turbo) (5.25 GB)

Qwen 3 4B Text Encoder FP4 Mixed (For Z Image Turbo) (3.24 GB)

They will be saved with SwarmUI expected file name so you don't need to modify presets

Don't forget to overwrite all previous files when updating, including utilities folder super important

Hopefully I am producing a new tutorial for NVFP4, quality comparison and LTX 2 presets coming soon

The Speed difference of NVFP4 is massive as below

7 January 2026 V120

SwarmUI tag error (visual not important) and Tecache import (ComfyUI broken it and i fixed the custom node myself) errors fixed

3 January 2026 V119

New Which_Bundles_Downloads_Which_Preset_Models.html added to the zip file

You can use that file to see which bundles in Windows_Start_Download_Models_App.bat downloads the necessary models for that preset

It covers all of the presets we have e.g. 2 examples shown below

1 January 2025 V118

I have spent literally like 1 day to compile new Qwen Image 2512 model into best quality Quant FP8 Scaled

I had to do more than 10 compiles, comparisons and find best workflow to compile

Each quality compile takes like 2 hours on RTX 5090

Our compiled model is significantly better than ComfyUI published raw FP8

You can compare and see yourself if you wish :)

This new model added into our downloader app as Qwen_Image_2512_Quant_FP8_Scaled.safetensors

Older Qwen Image removed from Qwen Image Core Bundle and Complete Image Generation and Editing Bundle and replaced with newer Qwen_Image_2512_Quant_FP8_Scaled.safetensors

With Amazing_SwarmUI_Presets_v43.json the following changes made - it took me literally like 1000 image generation and massive amount of analysis

Qwen Image 50 Steps Official Slow removed from presets since it is not useful or better than anything

Qwen Image Edit Plus 50 Steps removed from presets since it is not useful or better than anything

Qwen Image Realism Fast removed from presets since it is not useful or better than anything

Qwen Image UHD UlRealism 4+4 Steps - was typo duplicate removed

Those still exists in Amazing_SwarmUI_Presets_v41 file which is in older_presets folder

Use Qwen Image Core Bundle to get newer needed models

Qwen Image 8 Realism Fast upgraded to Qwen Image 8 Realism Fast 260101

Now uses Qwen Image 2512 and new turbo LoRA, upgrade difference is massive (same duration ultra fast) : https://imgsli.com/NDM3OTIz

Qwen Image 8 Steps Ultra Fast upgraded to Qwen Image 8 Steps Ultra Fast 260101

Now uses Qwen Image 2512 and new turbo LoRA - for non realistic images can be used

Qwen Image Edit Plus UHD Realism - 4+4 Steps upgraded into Qwen Image Edit 2511 UHD Realism - 4+4 Steps - 260101:

Now uses Qwen Image Edit 2511 and new turbo LoRA, upgrade difference is massive (same duration ultra fast) : https://imgsli.com/NDM3OTUx

Qwen Image 2512 UHD High Realism 4+4 Steps - 260101 added to the presets

Qwen Image UHD Realism Tier - 4+4 Steps upgraded to Qwen Image 2512 UHD Realism - 4+4 Steps - 260101

Now uses Qwen Image 2512 and new turbo LoRA, upgrade difference is massive (same duration ultra fast) : https://imgsli.com/NDM3OTUz

Both Qwen Images Stylized UHD Tier 1 and Qwen Images Stylized UHD Tier 2 are valid and up-to-date for Qwen Image Base or Qwen Image 2512 or Qwen Image Edit 2509 or Qwen Image Edit 2511.

Now I am adding last updated dates to titles and also preset descriptions so you can keep and not overwrite if you wish.

Hopefully I will convert all presets into ComfyUI workflows and publish them seperately in our ComfyUI installer.

31 December 2025 V116

After doing massive research on new FLUX 2 Turbo LoRA I have come up with the very best preset

The research literally took over 1000 image generation and lots of comparison

Recommended bundle is FLUX 2 Core Bundle but if your RAM becomes not enough there are 2 options

My models are default set 2048x2048 thus generated images are 4 megapixel and this model works best at this resolution

4 Megapixel image generation takes like 40 seconds on RTX 5090 with FP8 Scaled model

1st: Modify your backend and add this as args --cache-none

This really improves RAM management - VRAM is handled by ComfyUI so the RAM is the real issue for this model

Because it caches text encoder and completey deloads text encoder model - Mistral it is massive

2nd: Download and use FLUX 2 Low RAM Bundle

This bundle uses much lesser RAM and VRAM since it is Q4 GGUF both Text Encoder and the Model but lower quality and slower than FP8 Scaled

4 New presets added as below

FLUX 2 - 8 Steps Fast HQ - 251230, FLUX 2 - 8 Steps Fast HQ - Low RAM - 251230, FLUX 2 - 8 Steps Fast Realism - 251230, FLUX 2 - 8 Steps Fast Realism - Low RAM - 251230

Since requested now I will put date to new presets e.g. 251230 = 2025, month 12, day 30

Some download bugs fixed and downloader app improved

28 December 2025 V115

Z-Image-Turbo-Fun-Controlnet-Union-2.1-8steps (6.25 GB) added to the downloads

I believe it is a little bit better than previous Z-Image-Turbo-Fun-Controlnet-Union

To learn how to use ControlNet here : https://youtu.be/ezD6QO14kRc

24 December 2025 V114

Our downloader app now has Qwen_Image_Edit_2511_Quant_Scaled_FP8

I have made this model myself by using the special ComfyUI-QuantOps repo

It took like 2 hours on RTX 5090 and AMD 9950X CPU

This is the ultimate quality of quantization right now, half space, almost full BF16 quality, ultra fast

For this to work, we have to install this ComfyUI-QuantOps repo as a custom node into ComfyUI

Thus get the latest ComfyUI zip file, extract into your ComfyUI folder, overwrite and run Windows_Install_Or_Update_ComfyUI.bat file : https://www.patreon.com/posts/105023709

Moreover, Complete Image Generation and Editing Bundle and Qwen Image Core Bundle now downloads Qwen Image 2511 instead of older 2509

Futhermore, older Qwen Image 25-09 preset removed and new Qwen Image Edit - 2511 - 12 Steps: added

For this preset to work you need to have downloaded Qwen_Image_Edit_2511_Quant_Scaled_FP8 base model or BF16 version and Qwen-Image-FP8-Lightning-4steps-V1.0-fp32 LoRA - yes this LoRA is working best among all, i tested all

I recommend you to download Qwen Image Core Bundle

Qwen Image Edit 2511 Quality is significntly better than 25-09, you can see how 25-11 fixed image coloring issue here : https://imgsli.com/NDM2MzE3/0/1

It can do higher resolution as well so you can change base resolution of model to 1920x1920 from 1328x1328 for higher quality. 2560x2560 works even better compare for your use cases.

Qwen Image Edit 25-11 BF16 added to Qwen Image Edit models bundle

New tutorial Qwen Image Edit 25-09 vs 25-11 data shared with prompts and metadata having images and source images : Qwen_Image_Edit_2509_vs_2511.zip

Moreover now we dont enable Smart Image Prompt Resizing and preset will disable it auto for you so it is 1-click to use as usual

18 December 2025 V111

Wan 2.2 Core 8 Steps Bundle renamed into Wan 2.2 Core 4 Steps Bundle since now presets are using 4-steps with newest speed up LoRAs and models

2 New models added to following bundles (latest Wan 2.2 Text To Video speed up LoRAs)

Wan2.2_T2V_High_Noise_Lightx2v_4steps_LoRA_1217.safetensors (0.57 GB)

Wan2.2_T2V_Low_Noise_Lightx2v_4steps_LoRA_1217.safetensors (0.57 GB)

Bundles : Wan 2.2 Core 4 Steps Bundle and Wan 2.2 LoRAs and Complete Image Generation and Editing Bundle

Wan 2.2 presets updated to generate 16 FPS instead of 24 or 20 and 81 frames as default (5 seconds)

Moreover, we have enabled RIFE 2x FPS since our installers now auto installs RIFE thus you will get 5 seconds 32 FPS videos

Read the new presets descriptions and use Amazing_SwarmUI_Presets_v38.json

11 December 2025 V109

Model downloader app installers upgraded to Gradio 6.1.0

There was a massive bug that caused massive slowness for Downloader App and after I have reported they fixed the issue with Gradio 6.1.0 and now Model Downloader app is 100x more faster and responsive than before

Just extract zip file and overwrite older files and start model downloader as usual with Windows_Start_Download_Models_App.bat - it will auto upgrade and use latest version

3 December 2025 V108

Z-Image-Turbo-Fun-Controlnet-Union model added to the downloader

Included in Z Image Turbo Core Bundle, Complete Image Generation and Editing Bundle and Z Image Turbo Models

I did set its metadata accurately so it will work right away out of the box

SwarmUI and ComfyUI uses different folders so check ComfyUI checkbox if you need

Read its description for how to use but I am working on a tutorial to show Z Image Turbo models training in AI Toolkit + ControlNet + how to use it overal hopefully soon

Windows_Install_SwarmUI.bat and Windows_Update_SwarmUI.bat now includes ControlNet Preprocessor node of SwarmUI so when you run Update or make a fresh install it will auto get installed - this has so many different ControlNet processors

Alternatively on the interface you can use install button

This is included in RunPod installer as well

For MassedCompute SECourses image i emailed them to auto include hopefully

To be able to use newest ControlNet run Windows_Install_Or_Update_ComfyUI.bat first in your ComfyUI install and then use Windows_Update_SwarmUI.bat and then download the new model

1 December 2025 V107

Now when there is a SHA256 cache mismatch it will get latest SHA256 and fix the issue automatically and if needed it will redownload model

Wan2.2-I2V-A14B-Moe-Distill-Lightx2v-High_fp8_scaled.safetensors updated with better newer version - just redownload

2 New Wan 2.2 Text to Image Lightning LoRAs added

They really improving quality really cool improvement for free

They are added under Wan 2.2 LoRAs and Wan 2.2 Core 8 Steps Bundle

Wan2.2-T2V-A14B-4steps-lora-rank64-Seko-V2.0-Low.safetensors (1.14 GB)

Wan2.2-T2V-A14B-4steps-lora-rank64-Seko-V2.0-High.safetensors (1.14 GB)

So download them and then import new Amazing_SwarmUI_Presets_v36.json either manually or use Windows_Preset_Delete_Import.bat (deletes all and makes clean preset install recommended)

Then use Wan 22 Text To Video 8 Steps for these new LoRAs having preset

Currently this is the very best Lightning (4-steps) Wan 2.2 Text to Video workflow until a better one arrives hopefully

Also Z-Image Turbo model with Ostris AI Toolkit best workflow research is almost done will publish hopefully tomorrow

28 November 2025 V105

FLUX 2 presets added into Amazing_SwarmUI_Presets_v35.json

Z Image Turbo presets added into Amazing_SwarmUI_Presets_v35.json

Downloader app completey revamped and extremely improved

The initial loading of the app will take a while but after that you will see it is amazing

The model downloader app version increased to Gradio 6.0.1 with latest features

FLUX 2 models download added into downloader app as FLUX 2 Models and FLUX 2 Core Bundle

FLUX 2 models added into Complete Image Generation and Editing Bundle as well

Z Image Turbo models download added into downloader app as Z Image Turbo Models and Z Image Turbo Core Bundle

I also set the default resolution of the models as 1536x1536 pixels

I have added Z Image FP8 Convert into our SECourses Premium Musubi Tuner app and thus added FP8 Scaled version of the model as well into our downloader app

First time done by me so only SECourses followers getting it

FP8 Scaled version quality almost same as BF16 and GGUF Q8 but much faster than GGUF Q8 with only 6 GB size so lower VRAM GPUs can use it

You can see full size screenshot of the app here

Part 1 : https://huggingface.co/MonsterMMORPG/Wan_GGUF/resolve/main/real_app_ss_1.jpg

Part 2 : https://huggingface.co/MonsterMMORPG/Wan_GGUF/resolve/main/real_app_ss_2.jpg

SDXL Realism preset updated / upgraded

18 November 2025 V101

Wan 2.2 T2V High Noise 14B BF16 (26.61 GB) and Wan 2.2 T2V Low Noise 14B BF16 (26.61 GB) added to the Wan models list

BF16 works higher quality than FP16 in most cases

3 New Qwen Image Edit LoRAs added

Check their description to see example prompts

Qwen-Image-Edit-2509-Multiple-Angles-LoRA.safetensors (0.22 GB)

Qwen-Image-Edit-2509-Relight-LoRA.safetensors (0.22 GB)

Qwen-Image-Edit-2509-Fusion-LoRA.safetensors (0.22 GB)

BiRefNet HR (High Resolution) Background Remover https://www.patreon.com/posts/121679760

14 November 2025 V99

New SwarmUI bundle : Trained Models Min Requirements (Total: 31.13 GB, 11 models)

This model downloads all automatically downloaded models by SwarmUI for FLUX, Qwen, Wan, HiDream models including 2 LoRAs we use for Qwen

Super useful to use on RunPod and Massed Compute or at fresh installs

10 November 2025 V97

2 New Qwen Image models Realism LoRAs added to the model downloader

They are included in both Image Generation and Editing Bundle and Qwen Image Core Bundle

You can find them individually under Qwen Image Models

- Qwen_LoRA_Amateur_Photo_v1.safetensors and - Qwen_LoRA_Skin_Fix_v2.safetensors

Presets updated and 4 new presets added

Inpainting Preset - Extremely useful to inpaint your existing images and fix errors, improve face etc

Follow this tutorial video to understand how it works : Link will be added hopefully

Qwen Image UHD UlRealism 4+4 Steps : This is ultra realistic but works only for real life like images, so if you are generating an image with riding a Dragon won't work

You can increase/decrease selected LoRA weight according to your needs

Qwen Image UHD High Realism 4+4 Steps : This is a very balanced LoRA that brings both realism + keeps amazing capability of the model

Currently LoRA is set to 0.6 weight for self trained models

If you use on raw Qwen Model, make weight like 1.2 makes really good realism improvement

You can increase/decrease selected LoRA weight according to your needs

Outpainting Preset FLUX Dev Fill : This preset used FLUX Dev Fill model to outpaint images as shown in tutorial

FLUX DEV Fill model is now included in Image Generation and Editing Bundle since it is useful to fix mistakes or like remove text on existing images

Follow this tutorial video to understand how it works : Link will be added hopefully

30 October 2025 V96

Expired download token refreshed

New 2 presets added

Qwen Images Stylized UHD Tier 1 and Qwen Images Stylized UHD Tier 2

They are useful when you generate stylized images, style images after training

22 October 2025 V92

Presets and bundle downloads updated for Qwen Image Realism

Now on your self trained Qwen Image Models realism is next level here below few examples

https://www.reddit.com/r/SECourses/comments/1ocrv0x/qwen_image_training_almost_completed_here_not/

Import latest v29 preset file

You can use below presets for such realism

Qwen Image UHD Realism Tier 1 - 8+8 Steps

Qwen Image UHD Realism Tier 2 - 4+4 Steps

Qwen Image Edit Plus UHD Realism - 4+4 Steps

This is for training your subject on Qwen Imaged Edit Plus 2509 model with pure black control images - our Musubi Tuner already have this feature to generate such black control images

17 October 2025 V90

Missing LoRA added to bundles

2 LoRAs were downloaded into diffusion_models folder inaccurately and fixed this issue

Few downloaded model names fixed to match presets

Qwen_Image_FP8_Scaled.safetensors download error fixed

There is now 2 presets

Wan 22 Image To Video 4 Steps HQ

Wan 22 Image To Video 8 Steps UHQ

I prefer 8 steps but 4 is also great

15 October 2025 V86

Image Generation and Editing Bundle and Qwen Image Core Bundle updated

Qwen Image Edit Plus 2509 is now FP8_Scaled instead of GGUF Q8

Advantage is up to 1.5x more speed on modern GPUs, same quality, lesser VRAM

I did a converter script to convert this model into scaled FP8 and it took huge time

Scaled FP8 is many times more quality than just base FP8

Qwen Image Model updated to Qwen Image FP8 Scaled from GGUF Q8

New Wan 2.2 Image to Video base models added to the Wan 2.2 Core 8 Steps Bundle

I have converted these new models into FP8 Scaled myself

Their names are Wan2.2-I2V-A14B-Moe-Distill-Lightx2v-Low_fp8_scaled and Wan2.2-I2V-A14B-Moe-Distill-Lightx2v-High_fp8_scaled

These models are working without LoRAs with just 4 steps - and quality is amazing

They were published today :)

According to these new models the following presets updated

Qwen Image 50 Steps Official Slow

Qwen Image 8 Realism Fast

Qwen Image 8 Steps Ultra Fast

Qwen Image Edit Plus 12 Steps

Qwen Image Edit Plus 50 Steps

Qwen Image Realism Fast

New preset Wan 22 Image To Video 4 Steps HQ added and I recommend this for image to video generation now - best one and only 4 steps ultra fast

So download these updated bundles, don't worry it won't redownload existing models just new ones

Run Download Qwen Image Core Bundle and Download Wan 2.2 Core 8 Steps Bundle to new models

13 October 2025 V84

This is a major update for presets

With ComfyUI installer v56 now we have the missing new Samplers and Schedulers inside SwarmUI such as beta57, bong_tangent, res_2s : https://www.patreon.com/posts/105023709

So make sure to get latest ComfyUI installer zip file and run installer 1 time to get this update

I have updated images of all presets so that it will be way easier to recognize presets

New preset Qwen Image 8 Realism Fast added and it is really really much more realistic

I will hopefully update Qwen LoRA training post and add realistic results soon : https://www.patreon.com/posts/137551634

Of course it is realistic for other LoRAs or no LoRAs too

I did a massive research on Wan 2.2 image generation and we completey remade Wan 22 Image Realism preset

Now it is really really realistic look at this example : https://huggingface.co/MonsterMMORPG/Generative-AI/resolve/main/Wan_2_2.png

This also had our 2x upscale preset applied

For new Wan 2.2 and Qwen Image Realism presets to work you have to update your ComfyUI installation as mentioned above

For this update just import Amazing_SwarmUI_Presets_v24.json and overwrite older ones

26 September 2026 V82

Amazing ULR Downloader implemented into the app that lets you download from CivitAI and Hugging Face into relative or absolute path with custom file name support

It uses 16 connections and you can blazing fast download from even CivitAI

25 September 2026 V81

Qwen Image Lightning 8steps-V2 LoRA added to the downloader

This is a significant quality improvement compared to V1 and presets updated to use this LoRA

The bundles are also updated to download this new LoRA

FLUX SRPO models added to the downloader

Both BF16 and GGUF are available

FLUX SRPO is extremely realistic model

FLUX Bundle will now download FLUX SRPO BF16 model as well

I have uploaded a metadata fixed, so it will be auto set for you

Fixed metadata of FLUX Kontext Dev model as well with re-upload

Always check metadata of models in SwarmUI and verify accurate

FLUX Dev Models preset info and thumbnail image updated

FLUX Krea Dev Official removed since it is identical of FLUX Dev Models preset

FLUX Dev Models preset supports all FLUX Dev, FLUX Krea and FLUX SRPO

Just change base model

FLUX Kontext Dev Edit Images preset info, thumbnail image and settings updated

Qwen Image Edit Plus models added to the downloader and older Qwen Image Edit model replaced with Plus model in bundle (Qwen Image Edit 2509)

Presets are also updated to use this model

I have personally fixed metadata of the BF16 and FP8 version of the model

For GGUF again make sure that SwarmUI sees as Qwen Image Edit Plus the architecture

I added a json file for bundle downloaded GGUF Q8 so it will auto recognize it

Qwen Image Edit plus supports up to 3 input images so no more image sticthing is necessary and this model has huge quality

We are using Qwen Image Lightning 8steps-V2 LoRA with this model as well

If any better speed up LoRA arrives I will hopefully update

GGUF models are way slower compared to BF16 models on many hardware so if your VRAM is sufficient, prefer BF16 like on RTX A6000 GPUs

If your VRAM is not sufficient still you can compare speed of BF16 vs GGUF and see which one works best on your PC

ComfyUI auto does block swapping so you can even run BF16 on 24 GB GPUs

Make sure you have lots of RAM otherwise block swap will fail and your PC will freeze or use lower quant Like GGUF_8 or GGUF_5 etc

Under Other Models (e.g. Yolo Face Segment, Image Upscaling) section, Download All Auto Yolo Masking/Segment Models updated

Now it has Yolo V12 Large model as well : yolov12l-face.pt

Previously this was not working but after I opened an issue on SwarmUI and we did debug, it is fixed and now working

Face segment bundle is included in Qwen Image Core Bundle, FLUX Models Bundle and HiDream-I1 Dev Bundle

Example usage : photo of a man a man

Model downloader completeyt refactored for future easier development

Some of the files moved inside utilities folder and content of zip file simplified

Now we are not using Hugging Face downloader anymore and we are using much more robust uGet style 16 connection downloader that I developed

The advantage of this approach is that, it is extremely robust and works perfect even on low speed connections

It has 100% resume capability even if at stops at 99.99% percent

It auto checks SHA256 to verify accuracy of the download so the model is never corrupted

Sometimes you may get errors like on this on SwarmUI : https://github.com/mcmonkeyprojects/SwarmUI/issues/1065

On that case close SwarmUI, run Manually_Rebuild_SwarmUI_On_Errors.bat and restart SwarmUI

This will SwarmUI to compile necessary libraries on next run again

Use Windows_Preset_Delete_Import.bat to delete your all previous presets and import new updated ones : Amazing_SwarmUI_Presets_v22.json

Hopefully I will look for newer Wan 2.2 and Wan 2.1 speed up LoRAs to update presets even better quality

I recommend using 12 Steps Qwen Image Edit Plus and 8 Steps Qwen Image presets rather than 50 steps ones

Tutorials

Main ComfyUI + SwarmUI setup tutorial (4 May 2025) : https://youtu.be/fTzlQ0tjxj0

Wan 2.1 Tutorial shows how to use LoRA as well (19 May 2025) : https://youtu.be/XNcn845UXdw

ComfyUI + SwarmUI RunPod master tutorial (10 June 2025) : https://youtu.be/R02kPf9Y3_w

SwarmUI Wan 2.1 FusionX + Image Upscale tutorial (18 June 2025) : https://youtu.be/Xbn93GRQKsQ

SwarmUI FLUX Kontext (Ultimate image editing with just prompts) tutorial (27 June 2025) : https://youtu.be/adF9X9E0Chs

MultiTalk (image + audio = full animated lip synched video) main tutorial (ComfyUI only) (10 July 2025) : https://youtu.be/8cMIwS9qo4M

MultiTalk new workflows tutorial (ComfyUI only) (12 July 2025) : https://youtu.be/wgCtUeog41g

Wan 2.2 & FLUX Krea Full Tutorial - Automated Install - Ready Perfect Presets - SwarmUI with ComfyUI (2 August 2025) : https://youtu.be/8MvvuX4YPeo

Qwen Image Dominates Text-to-Image: 700+ Tests Reveal Why It's Better Than FLUX - Presets Published (7 August 2025) : https://youtu.be/R6h02YY6gUs

Wan 2.2, FLUX & Qwen Image Upgraded: Ultimate Tutorial for Open Source SOTA Image & Video Gen Models (19 August 2025) : https://youtu.be/3BFDcO2Ysu4

Qwen Image Edit Full Tutorial: 26 Different Demo Cases, Prompts & Images, Pwns FLUX Kontext Dev (23 August 2025) : https://youtu.be/gLCMhbsICEQ

Windows Requirements

Python 3.10, FFmpeg, CUDA 12.8, cuDNN 9.7 or above, C++ tools and Git

If it doesn't work make sure to below tutorial and install everything exactly as shown in this below tutorial

https://youtu.be/DrhUHnYfwC0

Requirements are important to have CUDA, C++ Tools, MSVC

Massed Compute (Recommend Cloud) :

Please register via this link : https://vm.massedcompute.com/signup?linkId=lp_034338&sourceId=secourses&tenantId=massed-compute

Use our coupon SECourses

Our coupon works on all GPUs now

H100 has amazing price and speed but you can use like RTX A6000 ADA as well

Full details here : https://www.patreon.com/posts/26671823

Then select our image SECourses from Creator dropdown

Then follow Massed_Compute_Instructions_READ.txt

Same as my any other Massed Compute installer script

Example tutorial for learn how to install and use Massed Compute

(Starts at 12:58) : https://youtu.be/KW-MHmoNcqo?si=G1WbG-Qw4ujWvOtG&t=778

RunPod (Cloud):

Please register via this link : https://runpod.io?ref=1aka98lq

Then follow Runpod_Instructions_READ.txt

Same as my any other RunPod installer script

Use the template written in Runpod_Instructions_READ.txt file

Example tutorial for learn how to install and use RunPod

(starts at 22:03) : https://youtu.be/KW-MHmoNcqo?si=QN8X8Sjn13ZYu-EU&t=1323

Total Models size you can download is over 3200 GB atm

OLDER VERSION HISTORY

12 September 2025 V80

The unified AI models downloader now supports newest Stable Diffusion WebUI Forge - Classic : https://www.patreon.com/posts/138680643

The unified AI models downloader now supports newest Stable Diffusion WebUI Forge - NEO : https://www.patreon.com/posts/138694680

Just check the checkbox for these 2 apps:

Forge WebUI / Automatic1111 Folder Structure

Example path : E:\Forge_Neo_v1\sd-webui-forge-classic\models

27 August 2025 V79

Demo images zip file upgraded to V3 since we published Nano Banana free image editing tutorial and compared it with Qwen Image Edit model for every case

Qwen_Edit_Demo_Images_With_Metadata_And_Prompts_v3.zip

Full tutorial link : https://youtu.be/qPUreQxB8zQ

A LoRA selection error of Qwen Image Edit model preset fixed with Amazing_SwarmUI_Presets_v21.json

22 August 2025 V77

Image upscale preset unified into single one

First use your regular preset then plus apply it

Prompt file modified to FLUX_Kontext_Qwen_Edit_Prompts.txt

Contains 26 unique cases and total 40 prompts

Demo images and results added, results have metadata so you can see and replicate

If Sage Attention on vs off, produces different results

Hugging Face download errors fixed

20 August 2025 V75

Use latest zip file extract and overwrite

Don't forget to edit metadata of the Qwen Image Edit model and set architecture of the model as Qwen Image Edit since it is not auto recognized yet

I have added 2 new presets called as Qwen Image Edit 50 Steps and Qwen Image Fast 12 Steps

Qwen Image Fast 12 Steps takes 30 seconds on RTX 5090

Qwen Image Edit 50 Steps takes 128 seconds on RTX 5090

Here a comparison of org image vs 12 steps vs 50 steps : https://imgsli.com/NDA4MTU1

Prompt is : change hair color to blue

The presets are using Qwen Image Edit model GGUF Q8 but you can use lower GGUF too

Our downloader has BF16, FP8, GGUF Q8, Q6, Q5, Q4

First update your ComfyUI and SwarmUI to latest version as shown in this tutorial and import presets and apply them : https://youtu.be/3BFDcO2Ysu4

This model doesn't work with input image but rather it takes input image from prompt like in FLUX Kontext : https://youtu.be/adF9X9E0Chs

This model doesn't support inpainting

Hopefully I will make a new tutorial for this but even ComfyUI official workflow is so bad atm - our preset way better

So we need some more time for this model to become better

Our Windows_Start_Download_Models_App.bat now have Qwen Image Edit model and it is also included in SwarmUI Qwen Image core bundle

SwarmUI update file will now auto install RIFE frame interpolation and Teacache

So just download bundle to download this model - it won't redownload existing models

19 August 2025 V72

Use latest zip file extract and overwrite

Full how to use tutorial published : https://youtu.be/3BFDcO2Ysu4

New following presets added and now we have Windows_Preset_Delete_Import.bat file

Execute this file while SwarmUI running, it will backup your existing presets into presets_backups folder, then delete all your presets and import newest presets

The added presets are

Qwen Image 8 Steps Ultra Fast 2x Upscale : Really upscales amazingly Qwen Image generations and fast too

You can increase steps count to 20 if not good enough

Wan 22 High Quality I2V 20 Steps : Super high quality but slow

Wan 22 High Quality T2V 20 Steps: Super high quality but slow

Wan 22 Image To Video 8 Steps: Really decent quality and ultra fast

Wan 22 Text To Video 8 Steps: Really decent quality and ultra fast

Wan 22 Image Realism: Super high quality and fast

This preset generates static images unlike videos

Wan 2.2 Image Generation Test Grids

Dragon Test Grid , Human Test Grid

Preparing these very best presets literally took days and 100s of generations and comparisons

To not delete your older presets, use import and import v15 and overwrite existing, it will only overwrite same name presets

On RunPod and Massed Compute use regular import preset feature same as before

Also use model downloader file and use Wan 2.2 Core 8 Steps Bundle (Total: 70.41 GB, 11 models) to download new 4 LoRAs

It will download new lightx2v Wan2.2-Lightning LoRAs

Some bundle downloads were broken for some LoRAs and these errors fixed

Windows SwarmUI installer will now auto install RIFE Frame Interpolation and Teacache automatically

13 August 2025 V66

Download newest v66 zip file, extract and overwrite previous files as usual

New amazing Qwen Image preset added : Qwen Image 8 Steps Ultra Fast

Literally 6x faster than Qwen Image High Quality preset

To use this preset use latest models downloader and download Qwen Image Core Bundle (Total: 30.83 GB, 4 models) or Qwen Image Lightning 8steps V1.1 LoRA (Fast inference LoRA) (1.58 GB)

If you had downloaded older bundle before, it will just download new LoRA

Moreover, import new preset json file and overwrite : Amazing_SwarmUI_Presets_v12.json

All Preset images updated with Qwen Image 8 Steps Ultra Fast preset - so now all preset thumbnails have text to easier recognize

All presets will automatically select accurate models based on our model downloader

I have tested all presets and verified all works

Sage Attention is now working with Qwen Image again - make sure to update your ComfyUI and SwarmUI to the latest version

FLUX Dev Official 2x Latent Upscale preset works with FLUX Krea Dev model as well just change base model

I also tested all best schedulers for FLUX Dev model and updated presets for best one

You can see grid comparison here : click to download

I also have tested new Scheduler named as KL Optimal (Nvidia AYS) with all samplers and it produces some realistic and

I think our very best preset still overall better but it is a good one to test in some cases

You can see its comparison grid here : click to download

Hopefully I will update Wan 2.2 presets with new LoRAs after testing and verifying they are better

8 August 2025 V63

Now supports Automatic1111 Web UI and SD Forge Web UI model structure too

Just select Forge WebUI / Automatic1111 Folder Structure checkbox and give model path like below

Windows e.g. : E:\Forge_Installer_v10\stable-diffusion-webui-forge\models

Massed Compute e.g. : /home/Ubuntu/Downloads/Forge_Installer_v10/stable-diffusion-webui-forge/models/,

RunPod e.g. : /workspace/stable-diffusion-webui-forge/models

New tutorial published : Qwen Image Dominates Text-to-Image: 700+ Tests Reveal Why It's Better Than FLUX - Presets Published > https://youtu.be/R6h02YY6gUs

SD Forge Web UI Installers updated for Windows, RunPod and Massed Compute and now supports RTX 5000 series too and working perfect with more robust, easier and faster install : https://www.patreon.com/posts/118442039

The error we were getting when downloading upscale models and face restoration models fixed : exists and seems populated. Skipping snapshot download for

6 August 2025 why Qwen Image is the New King and the Research

You can download all the grid tests (700+ generations) i did for Qwen Image inference parameter research here : https://huggingface.co/BestModelsv2/test/resolve/main/qwen_image_inference_research_grids.zip

Put the grids into this folder SwarmUI\Output\local\Grids and restart SwarmUI then you will be able to load and see each grid, tested parameters and results in full quality

Moroever I have compared FLUX Dev Official vs FLUX Krea Dev Official vs Qwen Image Realism Fast vs Qwen Image High Quality presets and you can see full quality result grid image here : https://huggingface.co/MonsterMMORPG/Generative-AI/resolve/main/FLUX_vs_Krea_vs_Qwen.png

Hopefully I will fully research Qwen Image training as i did for FLUX and then we will have a new amazing training workflow :

6 August 2025 V62

I have generated 600 images to find out the very best presets of Qwen Image model and now 2 new presets added to our presets with Amazing_SwarmUI_Presets_v9 file inside zip file

We have Qwen Image High Quality which generates highest quality overal with maximum prompt following - default negative prompt also set which improves quality further

And there is Qwen Image Realism Fast which is 2x faster since cfg is set to 1

This preset may require more generation to get best result but definitely better realism

Moreover downloader app now have the following models to download with 1 click

Under Image Generation Models under Qwen Image Models

Qwen_Image_Q4_1 (11.96 GB) - Qwen_Image_Q5_1 (14.33 GB) - Qwen_Image_Q6_K (15.67 GB) - Qwen_Image_Q8_0 (20.27 GB) - Qwen_Image_FP8_e4m3f (19.03 GB) - Qwen_Image_BF16 (38.05 GB)

Just use SwarmUI Bundles and then Qwen Image Core Bundle (Total: 29.24 GB, 3 models) and the presets will work right away - uses Q8

Or manually select your model

Q4_1 quality is also amazing so if you are on low VRAM you can use it

Still all models should work if you have sufficient RAM since SwarmUI uses ComfyUI which does auto block swapping

Hopefully tomorrow I will share all grids so that you can download and analyze yourself locally with SwarmUI

Sage attention may cause black output so until fixed disable if you get

AllowGpuSpecificOptimizations inside Server > Server Configuration may cause black output so disable until it is fixed

Update your both ComfyUI and SwarmUI to the latest version

This is clearly the new King of models better than FLUX and hopefully I will cover training fully soon after full research

Image resolution has to be divisible to 16 at the moment so if you get error pay attention to that

2 August 2025 V61

Wan 2.2 + FLUX Krea full tutorial published > https://youtu.be/8MvvuX4YPeo

After carefully testing more, bad results yielding presets removed

Thus I recommend delete all previous presets and import new Amazing_SwarmUI_Presets_v8 - screenshot

I also removed the following bundles since they were not necessary anymore with recent LoRAs + base Wan models

Wan 2.1 FusionX FP16 Phantom Bundle

Wan 2.1 FusionX FP8 Phantom Bundle

Moreover, all presets will automatically select models downloaded from our downloader app bundles now so that you won't be needed to manually select presets

The bundles you should download for SwarmUI are as below

Wan 2.2 Core 8 Steps Bundle (Total: 65.84 GB, 7 models) - FP8 Scaled base models + LoRA

Base models can be replaced with other variants like GGUF Q8 or FP16 etc

Wan 2.1 Core Models Bundle (GGUF Q6_K + Best LoRAs) (Total: 36.89 GB, 10 models) - GGUF Q6 base models + LoRA

Base models can be replaced with other variants like GGUF Q8 or FP8 etc

FLUX Models Bundle (Total: 100.77 GB, 10 models)

Base models can be replaced with other variants like GGUF Q8 or FP8 etc

All models still remaining so just search the model name and download whichever you want

31 July 2025 V60

Amazing new official FLUX Krea DEV model published

I have added both FP16 and GGUF Q8, Q6, Q5 and Q4 models

Q8 almost same quality as FP16

It is way more realistic here first comparison test I made : https://www.reddit.com/r/SECourses/comments/1me4jkb/flux_krea_dev_is_really_realistic_improvement/

The download button is under Image Generation Models▼ FLUX Models▼ - at the very top

FLUX Models is now seperated into normal Models vs GGUF variants as 2 tabs

Search function significantly improved, just search krea and test

New Wan 2.2 text-to-video high quality preset added

Sadly none of the fast LoRAs of Wan 2.1 working for Wan 2.2 text-to-video greatly so it is best to not use lora and do at least 20 steps until a LoRA is published

Image to video works perfect with our Wan 2.2 bundle LoRA and preset and 8 steps

31 July 2025 V58

Both ComfyUI and SwarmUI zip files updated

Now they are more robust to ensure updates goes smooth

Make sure update your ComfyUI backend and SwarmUI

After doing literally 100s of grid Wan 2.2 generations I have prepared amazing 8 steps 2 presets

These presets are even selecting the accurate models and LoRA

First download new SwarmUI bundle : Wan 2.2 Core 8 Steps Bundle (Total: 65.84 GB, 7 models) - use Windows_Start_Download_Models_App.bat from latest zip file

Then import new preset Amazing_SwarmUI_Presets_v6.json

Then reset params to default and direct apply preset you want to use

For text to video model just type your prompt and also set your desired resolution and set your desired Text2Video Frames count - default 73 frames thus 3 seconds (24 fps), for 5 seconds video set 121 frames

For image to video select your input image, type your prompt, set target resolution according to image aspect ratio and set your Video Frames under Image to Video tab - default 73 frames thus 3 seconds (24 fps), for 5 seconds video set 121 frames

Thats all and hit generate and enjoy

I am recording a tutorial video for Wan 2.2 hopefully

Moreover we are using UniPC sampler and Eular Ancestral behaves significantly different you can compare both

I am also using --use-sage-attention which brings good speed up and working great

Also we have Video_Models_Prompt_Generate_Guide.txt file

You can use this file in free Gemini in Google Studio ai to generate amazing prompts, just upload file and write what you want to generate : https://aistudio.google.com/prompts/new_chat (free)

16 July 2025 V57

New 4-8 steps LoRA added for Wan 2.1 models

Wan 2.1 14B LightX2V CFG Step Distill LoRA V2 (T2V + I2V) (Rank 64) (0.69 GB)

This LoRA will be now downloaded with bundles

Works great for both image to video and text to video with Wan 2.1 Models

CausVid LoRA preset renamed to LightX2V

Do 4-10 steps, 8 is probably best quality / speed

SECourses: FLUX, Tutorials, Guides, Resources, Training, Scripts PATREON 32 favs
VIEWS1
FILES34 files
POSTEDJul 18, 2026
ARCHIVEDJun 10, 2026