Patreon exclusive posts index to find our scripts easily, Patreon scripts updates history to see which updates arrived to which scripts and amazing Patreon special generative scripts list that you can use in any of your task.
Join discord to get help, chat, discuss and also tell me your discord username to get your special rank : SECourses Discord
Please also Star, Watch and Fork our Stable Diffusion & Generative AI GitHub repository and join our Reddit subreddit and follow me on LinkedIn (my real profile)
=======
Latest installer zip file : NVFP4_Quantizer_v2.zip
NVFP4 models are 100+% faster than BF16 or FP8 or GGUF models.
However compiling them is not trivial task at all
This app is based on https://github.com/NVIDIA/Model-Optimizer
But I had to do massive fixes and improvements to make it work just with FLUX
Currently only FLUX models tested and verified and I have compiled NVFP4 for FLUX SRPO model
You can download FLUX SRPO with our Model downloader app : https://www.patreon.com/posts/114517862
I have used max Quantization Algorithm since svdquant was so slow and I had to spent over 24 hours on RTX PRO 6000 to constant test and compile to make it work
Currently compiling FLUX model requiring 48 GB GPU and I recommend SimplePod RTX PRO 6000 if you want to try
You can of course use RunPod or Massed Compute
Massed Compute H200 would work amazing
I plan to add Qwen 2512 hopefully later and compile a better quality FLUX SRPO with svdquant
Generated FLUX SRPO NVFP4 is only 6.32 GB so it is amazing for low VRAM GPUs
More about NVFP4 shown in this tutorial : https://youtu.be/yOj9PYq3XYM
15 January 2026 V1.2
New Mixed NVFP4 Quantization for FLUX implemented
Mixed NVFP4 FLUX SRPO model published : https://www.patreon.com/posts/114517862
Windows Requirements
Python 3.10.11 and Git
Follow this requirements tutorial video exactly : https://youtu.be/DrhUHnYfwC0
Follow its updated post with links and screenshots exactly : https://www.patreon.com/posts/click-to-open-post-used-in-tutorial-111553210
For RunPod and Massed Compute
Follow
RunPod_SimplePod_Install_NVFP4_Instructions.txt
Massed_Compute_Install_Instructions.txt
SIMPLEPOD CHEAPER AND FASTER THAN RUNPOD
Now we fully support SimplePod as well please use this link to register : https://simplepod.ai/ref?user=secourses
SimplePod is faster and cheaper than RunPod and works exactly same
E.g. RTX 5090 on RunPod is 0.89 USD per hour, on SimplePod it is 0.45$ per hour,
RTX PRO 6000 on RunPod is 1.84 USD per hour and on SimplePod it is 0.79 USD per hour
Please use this template on SimplePod : https://dash.simplepod.ai/account/explore/100/ref-secourses/
For permanent storage, generate it from Storage tab with any name and size you want and when selecting template with above link, click Edit and Use, select Persistence Volume and change mount point to /workspace
Up-to-date SimplePod tutorial starting from 21:51 : https://youtu.be/yOj9PYq3XYM?si=Z86wZZLBeYzWo1Qo&t=1311
As usual follow Massed_Compute_Instructions_READ.txt and Runpod_SimplePod_Instructions_Musubi_Trainer.txt to install and use and watch the tutorials