Previous Post

NVFP4 Model Converter Gradio App For Windows and Linux

Next Post
NVFP4 Model Converter Gradio App For Windows and Linux
1 / 2
DESCRIPTION

Patreon exclusive posts index to find our scripts easily, Patreon scripts updates history to see which updates arrived to which scripts and amazing Patreon special generative scripts list that you can use in any of your task.

Join discord to get help, chat, discuss and also tell me your discord username to get your special rank : SECourses Discord

Please also Star, Watch and Fork our Stable Diffusion & Generative AI  GitHub repository and join our Reddit subreddit and follow me on LinkedIn (my real profile)

=======

Latest installer zip file : NVFP4_Quantizer_v2.zip

NVFP4 models are 100+% faster than BF16 or FP8 or GGUF models.

However compiling them is not trivial task at all

This app is based on https://github.com/NVIDIA/Model-Optimizer

But I had to do massive fixes and improvements to make it work just with FLUX

Currently only FLUX models tested and verified and I have compiled NVFP4 for FLUX SRPO model

You can download FLUX SRPO with our Model downloader app : https://www.patreon.com/posts/114517862

I have used max Quantization Algorithm since svdquant was so slow and I had to spent over 24 hours on RTX PRO 6000 to constant test and compile to make it work

Currently compiling FLUX model requiring 48 GB GPU and I recommend SimplePod RTX PRO 6000 if you want to try

You can of course use RunPod or Massed Compute

Massed Compute H200 would work amazing

I plan to add Qwen 2512 hopefully later and compile a better quality FLUX SRPO with svdquant

Generated FLUX SRPO NVFP4 is only 6.32 GB so it is amazing for low VRAM GPUs

More about NVFP4 shown in this tutorial : https://youtu.be/yOj9PYq3XYM

15 January 2026 V1.2

New Mixed NVFP4 Quantization for FLUX implemented

Mixed NVFP4 FLUX SRPO model published : https://www.patreon.com/posts/114517862

https://imgsli.com/NDQyNjk5

Windows Requirements

Python 3.10.11 and Git

Follow this requirements tutorial video exactly : https://youtu.be/DrhUHnYfwC0

Follow its updated post with links and screenshots exactly : https://www.patreon.com/posts/click-to-open-post-used-in-tutorial-111553210

For RunPod and Massed Compute

Follow

RunPod_SimplePod_Install_NVFP4_Instructions.txt

Massed_Compute_Install_Instructions.txt

SIMPLEPOD CHEAPER AND FASTER THAN RUNPOD

Now we fully support SimplePod as well please use this link to register : https://simplepod.ai/ref?user=secourses

SimplePod is faster and cheaper than RunPod and works exactly same

E.g. RTX 5090 on RunPod is 0.89 USD per hour, on SimplePod it is 0.45$ per hour,

RTX PRO 6000 on RunPod is 1.84 USD per hour and on SimplePod it is 0.79 USD per hour

Please use this template on SimplePod : https://dash.simplepod.ai/account/explore/100/ref-secourses/

For permanent storage, generate it from Storage tab with any name and size you want and when selecting template with above link, click Edit and Use, select Persistence Volume and change mount point to /workspace

Up-to-date SimplePod tutorial starting from 21:51 : https://youtu.be/yOj9PYq3XYM?si=Z86wZZLBeYzWo1Qo&t=1311

As usual follow Massed_Compute_Instructions_READ.txt and Runpod_SimplePod_Instructions_Musubi_Trainer.txt to install and use and watch the tutorials

SECourses: FLUX, Tutorials, Guides, Resources, Training, Scripts PATREON 32 favs
VIEWS1
FILES3 files
POSTEDJan 14, 2026
ARCHIVEDJan 14, 2026