SwarmUI Auto Installer + The Ultimate Image and Video AI Models Downloader - For Windows, RunPod and Massed Compute - Ultimate Compilation
🕑 Added 2025-10-13 00:00:00 +0000 UTCUnified AI Models Downloader for SwarmUI, ComfyUI, Automatic1111 and Forge Web UI, Forge Web UI Classic and Neo - Supports SD 1.5, SDXL, FLUX, FLUX Krea, FLUX SRPO, Wan 2.1, Wan 2.2, Qwen Image, Qwen Image Edit Plus 2509, SwarmUI presets to generate images and videos with FLUX, Qwen Image, SDXL and Wan 2.1, Wan 2.2 Models and more
Patreon exclusive posts index to find our scripts easily, Patreon scripts updates history to see which updates arrived to which scripts and amazing Patreon special generative scripts list that you can use in any of your task.
Join discord to get help, chat, discuss and also tell me your discord username to get your special rank : SECourses Discord
Please also Star, Watch and Fork our Stable Diffusion & Generative AI GitHub repository and join our Reddit subreddit and follow me on LinkedIn (my real profile)
=======
Latest zip file : SwarmUI_Model_Downloader_v95.zip
ComfyUI Back-end Installer with Torch 2.8 with CUDA 12.9, Triton (3.4+) and DeepSpeed (0.16.4), insightface (0.7.3), onnxruntime-gpu and Flash Attention (2.8.3 - I compiled for all 4 Python versions), Sage Attention 2.2, xFormers into our VENV Windows : https://www.patreon.com/posts/105023709
Definitely use our ComfyUI solo backend installer otherwise some of the newest stuff may fail
Example existing ComfyUI backend : E:\Comfy_UI_V41\ComfyUI\main.py
Sage Attention is optional
If you use Sage Attention and get black output, enable Display Advanced Options of SwarmUI and go Advanced Sampling and make Preferred DType = Default (16 bit)
If your GPU starts using shared VRAM for any reason, add this command to make it avoid that --reserve-vram 3
This command will preserve 3 GB VRAM for other tasks
It is like --use-sage-attention command
27 October 2025 V95
Expired download token refreshed
22 October 2025 V92
Presets and bundle downloads updated for Qwen Image Realism
Now on your self trained Qwen Image Models realism is next level here below few examples
Import latest v29 preset file
You can use below presets for such realism
Qwen Image UHD Realism Tier 1 - 8+8 Steps
Qwen Image UHD Realism Tier 2 - 4+4 Steps
Qwen Image Edit Plus UHD Realism - 4+4 Steps
This is for training your subject on Qwen Imaged Edit Plus 2509 model with pure black control images - our Musubi Tuner already have this feature to generate such black control images

17 October 2025 V90
Missing LoRA added to bundles
2 LoRAs were downloaded into diffusion_models folder inaccurately and fixed this issue
Few downloaded model names fixed to match presets
Qwen_Image_FP8_Scaled.safetensors download error fixed
There is now 2 presets
Wan 22 Image To Video 4 Steps HQ
Wan 22 Image To Video 8 Steps UHQ
I prefer 8 steps but 4 is also great
15 October 2025 V86
Image Generation and Editing Bundle and Qwen Image Core Bundle updated
Qwen Image Edit Plus 2509 is now FP8_Scaled instead of GGUF Q8
Advantage is up to 1.5x more speed on modern GPUs, same quality, lesser VRAM
I did a converter script to convert this model into scaled FP8 and it took huge time
Scaled FP8 is many times more quality than just base FP8
Qwen Image Model updated to Qwen Image FP8 Scaled from GGUF Q8
New Wan 2.2 Image to Video base models added to the Wan 2.2 Core 8 Steps Bundle
I have converted these new models into FP8 Scaled myself
Their names are Wan2.2-I2V-A14B-Moe-Distill-Lightx2v-Low_fp8_scaled and Wan2.2-I2V-A14B-Moe-Distill-Lightx2v-High_fp8_scaled
These models are working without LoRAs with just 4 steps - and quality is amazing
They were published today :)
According to these new models the following presets updated
Qwen Image 50 Steps Official Slow
Qwen Image 8 Realism Fast
Qwen Image 8 Steps Ultra Fast
Qwen Image Edit Plus 12 Steps
Qwen Image Edit Plus 50 Steps
Qwen Image Realism Fast
New preset Wan 22 Image To Video 4 Steps HQ added and I recommend this for image to video generation now - best one and only 4 steps ultra fast
So download these updated bundles, don't worry it won't redownload existing models just new ones
Run Download Qwen Image Core Bundle and Download Wan 2.2 Core 8 Steps Bundle to new models
13 October 2025 V84
This is a major update for presets
With ComfyUI installer v56 now we have the missing new Samplers and Schedulers inside SwarmUI such as beta57, bong_tangent, res_2s : https://www.patreon.com/posts/105023709
So make sure to get latest ComfyUI installer zip file and run installer 1 time to get this update
I have updated images of all presets so that it will be way easier to recognize presets
New preset Qwen Image 8 Realism Fast added and it is really really much more realistic
I will hopefully update Qwen LoRA training post and add realistic results soon : https://www.patreon.com/posts/137551634
Of course it is realistic for other LoRAs or no LoRAs too
I did a massive research on Wan 2.2 image generation and we completey remade Wan 22 Image Realism preset
Now it is really really realistic look at this example : https://huggingface.co/MonsterMMORPG/Generative-AI/resolve/main/Wan_2_2.png
This also had our 2x upscale preset applied
For new Wan 2.2 and Qwen Image Realism presets to work you have to update your ComfyUI installation as mentioned above
For this update just import Amazing_SwarmUI_Presets_v24.json and overwrite older ones

26 September 2026 V82
Amazing ULR Downloader implemented into the app that lets you download from CivitAI and Hugging Face into relative or absolute path with custom file name support
It uses 16 connections and you can blazing fast download from even CivitAI

25 September 2026 V81
Qwen Image Lightning 8steps-V2 LoRA added to the downloader
This is a significant quality improvement compared to V1 and presets updated to use this LoRA
The bundles are also updated to download this new LoRA
FLUX SRPO models added to the downloader
Both BF16 and GGUF are available
FLUX SRPO is extremely realistic model
FLUX Bundle will now download FLUX SRPO BF16 model as well
I have uploaded a metadata fixed, so it will be auto set for you
Fixed metadata of FLUX Kontext Dev model as well with re-upload
Always check metadata of models in SwarmUI and verify accurate
FLUX Dev Models preset info and thumbnail image updated
FLUX Krea Dev Official removed since it is identical of FLUX Dev Models preset
FLUX Dev Models preset supports all FLUX Dev, FLUX Krea and FLUX SRPO
Just change base model
FLUX Kontext Dev Edit Images preset info, thumbnail image and settings updated
Qwen Image Edit Plus models added to the downloader and older Qwen Image Edit model replaced with Plus model in bundle (Qwen Image Edit 2509)
Presets are also updated to use this model
I have personally fixed metadata of the BF16 and FP8 version of the model
For GGUF again make sure that SwarmUI sees as Qwen Image Edit Plus the architecture
I added a json file for bundle downloaded GGUF Q8 so it will auto recognize it
Qwen Image Edit plus supports up to 3 input images so no more image sticthing is necessary and this model has huge quality
We are using Qwen Image Lightning 8steps-V2 LoRA with this model as well
If any better speed up LoRA arrives I will hopefully update
GGUF models are way slower compared to BF16 models on many hardware so if your VRAM is sufficient, prefer BF16 like on RTX A6000 GPUs
If your VRAM is not sufficient still you can compare speed of BF16 vs GGUF and see which one works best on your PC
ComfyUI auto does block swapping so you can even run BF16 on 24 GB GPUs
Make sure you have lots of RAM otherwise block swap will fail and your PC will freeze or use lower quant Like GGUF_8 or GGUF_5 etc
Under Other Models (e.g. Yolo Face Segment, Image Upscaling) section, Download All Auto Yolo Masking/Segment Models updated
Now it has Yolo V12 Large model as well : yolov12l-face.pt
Previously this was not working but after I opened an issue on SwarmUI and we did debug, it is fixed and now working
Face segment bundle is included in Qwen Image Core Bundle, FLUX Models Bundle and HiDream-I1 Dev Bundle
Example usage : photo of a man <segment:yolo-yolov12l-face.pt> a man
Model downloader completeyt refactored for future easier development
Some of the files moved inside utilities folder and content of zip file simplified
Now we are not using Hugging Face downloader anymore and we are using much more robust uGet style 16 connection downloader that I developed
The advantage of this approach is that, it is extremely robust and works perfect even on low speed connections
It has 100% resume capability even if at stops at 99.99% percent
It auto checks SHA256 to verify accuracy of the download so the model is never corrupted
Sometimes you may get errors like on this on SwarmUI : https://github.com/mcmonkeyprojects/SwarmUI/issues/1065
On that case close SwarmUI, run Manually_Rebuild_SwarmUI_On_Errors.bat and restart SwarmUI
This will SwarmUI to compile necessary libraries on next run again
Use Windows_Preset_Delete_Import.bat to delete your all previous presets and import new updated ones : Amazing_SwarmUI_Presets_v22.json
Hopefully I will look for newer Wan 2.2 and Wan 2.1 speed up LoRAs to update presets even better quality
I recommend using 12 Steps Qwen Image Edit Plus and 8 Steps Qwen Image presets rather than 50 steps ones
Tutorials
Main ComfyUI + SwarmUI setup tutorial (4 May 2025) : https://youtu.be/fTzlQ0tjxj0
Wan 2.1 Tutorial shows how to use LoRA as well (19 May 2025) : https://youtu.be/XNcn845UXdw
ComfyUI + SwarmUI RunPod master tutorial (10 June 2025) : https://youtu.be/R02kPf9Y3_w
SwarmUI Wan 2.1 FusionX + Image Upscale tutorial (18 June 2025) : https://youtu.be/Xbn93GRQKsQ
SwarmUI FLUX Kontext (Ultimate image editing with just prompts) tutorial (27 June 2025) : https://youtu.be/adF9X9E0Chs
MultiTalk (image + audio = full animated lip synched video) main tutorial (ComfyUI only) (10 July 2025) : https://youtu.be/8cMIwS9qo4M
MultiTalk new workflows tutorial (ComfyUI only) (12 July 2025) : https://youtu.be/wgCtUeog41g
Wan 2.2 & FLUX Krea Full Tutorial - Automated Install - Ready Perfect Presets - SwarmUI with ComfyUI (2 August 2025) : https://youtu.be/8MvvuX4YPeo
Qwen Image Dominates Text-to-Image: 700+ Tests Reveal Why It's Better Than FLUX - Presets Published (7 August 2025) : https://youtu.be/R6h02YY6gUs
Wan 2.2, FLUX & Qwen Image Upgraded: Ultimate Tutorial for Open Source SOTA Image & Video Gen Models (19 August 2025) : https://youtu.be/3BFDcO2Ysu4
Qwen Image Edit Full Tutorial: 26 Different Demo Cases, Prompts & Images, Pwns FLUX Kontext Dev (23 August 2025) : https://youtu.be/gLCMhbsICEQ
Windows Requirements
Python 3.10, FFmpeg, CUDA 12.8, cuDNN 9.7 or above, C++ tools and Git
If it doesn't work make sure to below tutorial and install everything exactly as shown in this below tutorial
Requirements are important to have CUDA, C++ Tools, MSVC
Massed Compute (Recommend Cloud) :
Please register via this link : https://vm.massedcompute.com/signup?linkId=lp_034338&sourceId=secourses&tenantId=massed-compute
Use our coupon SECourses
Our coupon works on all GPUs now
H100 has amazing price and speed but you can use like RTX A6000 ADA as well
Full details here : https://www.patreon.com/posts/26671823
Then select our image SECourses from Creator dropdown
Then follow Massed_Compute_Instructions_READ.txt
Same as my any other Massed Compute installer script
Example tutorial for learn how to install and use Massed Compute
(Starts at 12:58) : https://youtu.be/KW-MHmoNcqo?si=G1WbG-Qw4ujWvOtG&t=778
RunPod (Cloud):
Please register via this link : https://runpod.io?ref=1aka98lq
Then follow Runpod_Instructions_READ.txt
Same as my any other RunPod installer script
Use the template written in Runpod_Instructions_READ.txt file
Example tutorial for learn how to install and use RunPod
(starts at 22:03) : https://youtu.be/KW-MHmoNcqo?si=QN8X8Sjn13ZYu-EU&t=1323
Screenshots of Presets

Total Models size you can download is over 2700 GB atm

Screenshot of Upgraded Ultra Robust Model Downloader



OLDER VERSION HISTORY
12 September 2025 V80
The unified AI models downloader now supports newest Stable Diffusion WebUI Forge - Classic : https://www.patreon.com/posts/138680643
The unified AI models downloader now supports newest Stable Diffusion WebUI Forge - NEO : https://www.patreon.com/posts/138694680
Just check the checkbox for these 2 apps:
Forge WebUI / Automatic1111 Folder Structure
Example path : E:\Forge_Neo_v1\sd-webui-forge-classic\models
27 August 2025 V79
Demo images zip file upgraded to V3 since we published Nano Banana free image editing tutorial and compared it with Qwen Image Edit model for every case
Full tutorial link : https://youtu.be/qPUreQxB8zQ
A LoRA selection error of Qwen Image Edit model preset fixed with Amazing_SwarmUI_Presets_v21.json
22 August 2025 V77
Image upscale preset unified into single one
First use your regular preset then plus apply it
Prompt file modified to FLUX_Kontext_Qwen_Edit_Prompts.txt
Contains 26 unique cases and total 40 prompts
Demo images and results added, results have metadata so you can see and replicate
If Sage Attention on vs off, produces different results
Hugging Face download errors fixed
20 August 2025 V75
Use latest zip file extract and overwrite
Don't forget to edit metadata of the Qwen Image Edit model and set architecture of the model as Qwen Image Edit since it is not auto recognized yet
I have added 2 new presets called as Qwen Image Edit 50 Steps and Qwen Image Fast 12 Steps
Qwen Image Fast 12 Steps takes 30 seconds on RTX 5090
Qwen Image Edit 50 Steps takes 128 seconds on RTX 5090
Here a comparison of org image vs 12 steps vs 50 steps : https://imgsli.com/NDA4MTU1
Prompt is : change hair color to blue
The presets are using Qwen Image Edit model GGUF Q8 but you can use lower GGUF too
Our downloader has BF16, FP8, GGUF Q8, Q6, Q5, Q4
First update your ComfyUI and SwarmUI to latest version as shown in this tutorial and import presets and apply them : https://youtu.be/3BFDcO2Ysu4
This model doesn't work with input image but rather it takes input image from prompt like in FLUX Kontext : https://youtu.be/adF9X9E0Chs
This model doesn't support inpainting
Hopefully I will make a new tutorial for this but even ComfyUI official workflow is so bad atm - our preset way better
So we need some more time for this model to become better
Our Windows_Start_Download_Models_App.bat now have Qwen Image Edit model and it is also included in SwarmUI Qwen Image core bundle
SwarmUI update file will now auto install RIFE frame interpolation and Teacache
So just download bundle to download this model - it won't redownload existing models
19 August 2025 V72
Use latest zip file extract and overwrite
Full how to use tutorial published : https://youtu.be/3BFDcO2Ysu4
New following presets added and now we have Windows_Preset_Delete_Import.bat file
Execute this file while SwarmUI running, it will backup your existing presets into presets_backups folder, then delete all your presets and import newest presets
The added presets are
Qwen Image 8 Steps Ultra Fast 2x Upscale : Really upscales amazingly Qwen Image generations and fast too
You can increase steps count to 20 if not good enough
Wan 22 High Quality I2V 20 Steps : Super high quality but slow
Wan 22 High Quality T2V 20 Steps: Super high quality but slow
Wan 22 Image To Video 8 Steps: Really decent quality and ultra fast
Wan 22 Text To Video 8 Steps: Really decent quality and ultra fast
Wan 22 Image Realism: Super high quality and fast
This preset generates static images unlike videos
Wan 2.2 Image Generation Test Grids
Preparing these very best presets literally took days and 100s of generations and comparisons
To not delete your older presets, use import and import v15 and overwrite existing, it will only overwrite same name presets
On RunPod and Massed Compute use regular import preset feature same as before
Also use model downloader file and use Wan 2.2 Core 8 Steps Bundle (Total: 70.41 GB, 11 models) to download new 4 LoRAs
It will download new lightx2v Wan2.2-Lightning LoRAs
Some bundle downloads were broken for some LoRAs and these errors fixed
Windows SwarmUI installer will now auto install RIFE Frame Interpolation and Teacache automatically
13 August 2025 V66
Download newest v66 zip file, extract and overwrite previous files as usual
New amazing Qwen Image preset added : Qwen Image 8 Steps Ultra Fast
Literally 6x faster than Qwen Image High Quality preset
To use this preset use latest models downloader and download Qwen Image Core Bundle (Total: 30.83 GB, 4 models) or Qwen Image Lightning 8steps V1.1 LoRA (Fast inference LoRA) (1.58 GB)
If you had downloaded older bundle before, it will just download new LoRA
Moreover, import new preset json file and overwrite : Amazing_SwarmUI_Presets_v12.json
All Preset images updated with Qwen Image 8 Steps Ultra Fast preset - so now all preset thumbnails have text to easier recognize
All presets will automatically select accurate models based on our model downloader
I have tested all presets and verified all works
Sage Attention is now working with Qwen Image again - make sure to update your ComfyUI and SwarmUI to the latest version
FLUX Dev Official 2x Latent Upscale preset works with FLUX Krea Dev model as well just change base model
I also tested all best schedulers for FLUX Dev model and updated presets for best one
You can see grid comparison here : click to download
I also have tested new Scheduler named as KL Optimal (Nvidia AYS) with all samplers and it produces some realistic and
I think our very best preset still overall better but it is a good one to test in some cases
You can see its comparison grid here : click to download
Hopefully I will update Wan 2.2 presets with new LoRAs after testing and verifying they are better
8 August 2025 V63
Now supports Automatic1111 Web UI and SD Forge Web UI model structure too
Just select Forge WebUI / Automatic1111 Folder Structure checkbox and give model path like below
Windows e.g. : E:\Forge_Installer_v10\stable-diffusion-webui-forge\models
Massed Compute e.g. : /home/Ubuntu/Downloads/Forge_Installer_v10/stable-diffusion-webui-forge/models/,
RunPod e.g. : /workspace/stable-diffusion-webui-forge/models
New tutorial published : Qwen Image Dominates Text-to-Image: 700+ Tests Reveal Why It's Better Than FLUX - Presets Published > https://youtu.be/R6h02YY6gUs
SD Forge Web UI Installers updated for Windows, RunPod and Massed Compute and now supports RTX 5000 series too and working perfect with more robust, easier and faster install : https://www.patreon.com/posts/118442039
The error we were getting when downloading upscale models and face restoration models fixed : exists and seems populated. Skipping snapshot download for
6 August 2025 why Qwen Image is the New King and the Research
You can download all the grid tests (700+ generations) i did for Qwen Image inference parameter research here : https://huggingface.co/BestModelsv2/test/resolve/main/qwen_image_inference_research_grids.zip
Put the grids into this folder SwarmUI\Output\local\Grids and restart SwarmUI then you will be able to load and see each grid, tested parameters and results in full quality
Moroever I have compared FLUX Dev Official vs FLUX Krea Dev Official vs Qwen Image Realism Fast vs Qwen Image High Quality presets and you can see full quality result grid image here : https://huggingface.co/MonsterMMORPG/Generative-AI/resolve/main/FLUX_vs_Krea_vs_Qwen.png
Hopefully I will fully research Qwen Image training as i did for FLUX and then we will have a new amazing training workflow :
6 August 2025 V62
I have generated 600 images to find out the very best presets of Qwen Image model and now 2 new presets added to our presets with Amazing_SwarmUI_Presets_v9 file inside zip file
We have Qwen Image High Quality which generates highest quality overal with maximum prompt following - default negative prompt also set which improves quality further
And there is Qwen Image Realism Fast which is 2x faster since cfg is set to 1
This preset may require more generation to get best result but definitely better realism
Moreover downloader app now have the following models to download with 1 click
Under Image Generation Models under Qwen Image Models
Qwen_Image_Q4_1 (11.96 GB) - Qwen_Image_Q5_1 (14.33 GB) - Qwen_Image_Q6_K (15.67 GB) - Qwen_Image_Q8_0 (20.27 GB) - Qwen_Image_FP8_e4m3f (19.03 GB) - Qwen_Image_BF16 (38.05 GB)
Just use SwarmUI Bundles and then Qwen Image Core Bundle (Total: 29.24 GB, 3 models) and the presets will work right away - uses Q8
Or manually select your model
Q4_1 quality is also amazing so if you are on low VRAM you can use it
Still all models should work if you have sufficient RAM since SwarmUI uses ComfyUI which does auto block swapping
Hopefully tomorrow I will share all grids so that you can download and analyze yourself locally with SwarmUI
Sage attention may cause black output so until fixed disable if you get
AllowGpuSpecificOptimizations inside Server > Server Configuration may cause black output so disable until it is fixed
Update your both ComfyUI and SwarmUI to the latest version
This is clearly the new King of models better than FLUX and hopefully I will cover training fully soon after full research
Image resolution has to be divisible to 16 at the moment so if you get error pay attention to that
2 August 2025 V61
Wan 2.2 + FLUX Krea full tutorial published > https://youtu.be/8MvvuX4YPeo
After carefully testing more, bad results yielding presets removed
Thus I recommend delete all previous presets and import new Amazing_SwarmUI_Presets_v8 - screenshot
I also removed the following bundles since they were not necessary anymore with recent LoRAs + base Wan models
Wan 2.1 FusionX FP16 Phantom Bundle
Wan 2.1 FusionX FP8 Phantom Bundle
Moreover, all presets will automatically select models downloaded from our downloader app bundles now so that you won't be needed to manually select presets
The bundles you should download for SwarmUI are as below
Wan 2.2 Core 8 Steps Bundle (Total: 65.84 GB, 7 models) - FP8 Scaled base models + LoRA
Base models can be replaced with other variants like GGUF Q8 or FP16 etc
Wan 2.1 Core Models Bundle (GGUF Q6_K + Best LoRAs) (Total: 36.89 GB, 10 models) - GGUF Q6 base models + LoRA
Base models can be replaced with other variants like GGUF Q8 or FP8 etc
FLUX Models Bundle (Total: 100.77 GB, 10 models)
Base models can be replaced with other variants like GGUF Q8 or FP8 etc
All models still remaining so just search the model name and download whichever you want
31 July 2025 V60
Amazing new official FLUX Krea DEV model published
I have added both FP16 and GGUF Q8, Q6, Q5 and Q4 models
Q8 almost same quality as FP16
It is way more realistic here first comparison test I made : https://www.reddit.com/r/SECourses/comments/1me4jkb/flux_krea_dev_is_really_realistic_improvement/
The download button is under Image Generation Models▼ FLUX Models▼ - at the very top
FLUX Models is now seperated into normal Models vs GGUF variants as 2 tabs
Search function significantly improved, just search krea and test
New Wan 2.2 text-to-video high quality preset added
Sadly none of the fast LoRAs of Wan 2.1 working for Wan 2.2 text-to-video greatly so it is best to not use lora and do at least 20 steps until a LoRA is published
Image to video works perfect with our Wan 2.2 bundle LoRA and preset and 8 steps
31 July 2025 V58
Both ComfyUI and SwarmUI zip files updated
Now they are more robust to ensure updates goes smooth
Make sure update your ComfyUI backend and SwarmUI
After doing literally 100s of grid Wan 2.2 generations I have prepared amazing 8 steps 2 presets
These presets are even selecting the accurate models and LoRA
First download new SwarmUI bundle : Wan 2.2 Core 8 Steps Bundle (Total: 65.84 GB, 7 models) - use Windows_Start_Download_Models_App.bat from latest zip file
Then import new preset Amazing_SwarmUI_Presets_v6.json
Then reset params to default and direct apply preset you want to use
For text to video model just type your prompt and also set your desired resolution and set your desired Text2Video Frames count - default 73 frames thus 3 seconds (24 fps), for 5 seconds video set 121 frames
For image to video select your input image, type your prompt, set target resolution according to image aspect ratio and set your Video Frames under Image to Video tab - default 73 frames thus 3 seconds (24 fps), for 5 seconds video set 121 frames
Thats all and hit generate and enjoy
I am recording a tutorial video for Wan 2.2 hopefully
Moreover we are using UniPC sampler and Eular Ancestral behaves significantly different you can compare both
I am also using --use-sage-attention which brings good speed up and working great
Also we have Video_Models_Prompt_Generate_Guide.txt file
You can use this file in free Gemini in Google Studio ai to generate amazing prompts, just upload file and write what you want to generate : https://aistudio.google.com/prompts/new_chat (free)
16 July 2025 V57
New 4-8 steps LoRA added for Wan 2.1 models
Wan 2.1 14B LightX2V CFG Step Distill LoRA V2 (T2V + I2V) (Rank 64) (0.69 GB)
This LoRA will be now downloaded with bundles
Works great for both image to video and text to video with Wan 2.1 Models
CausVid LoRA preset renamed to LightX2V
Do 4-10 steps, 8 is probably best quality / speed
Comments
Furkan Gözükara
It can cause lesser performance and more vram
Furkan Gözükara
Yes please follow video you to install both comfyui and swarmui : https://youtu.be/c3gEoAyL2IE?si=F4I_8MK-Hq8xt8sl
simon DINEEN
Error] Final error (4) while initializing backend #4 - ComfyUI Self-Starting, giving up: System.ArgumentException: The value cannot be an empty string. (Parameter 'path') anything I need to be concerned about?
simon DINEEN
so what happens if I don't have flash attention?
Furkan Gözükara
yes i will make tomorrow hopefully. today recorded windows tutorial editing it.
Dan
I have bought credits via your link and tried several times to set up SwarmUI in massed compute (I love to try some high end GPUs), but this is still confusing and not working. Can you please create a Step by step tutorial how to use SwarmUI (with the Comfy backend) on massed compute, I can get this to work locally but not on Massed compute! PLEASE create a super simple tutorial for swarmui on massed. thank you
Furkan Gözükara
I still prefer 3.10.11. it works perfect and it has the widest node support etc. unless something forces and broken in 3.10.11 i prefer it. you are welcome thanks for comment
David
Furkan, it's always me, a new version come and I have to install it. I decided to install from zero 😅 Your ComfyUI backend and this Swarm UI, till now I used (as you told us) python 3.10.11 installed as Windows system environment. Now, you compiled it for 3.11, 3.12 and 3.13 too. So, what's best and faster? I have Windows 11, 5090 and 9950X3D. Thank you
Furkan Gözükara
just get latest zip file v82, extract and overwrite and you will see v82 now. fixed that version mismatch
ranjeet
I already have v80 of swarmUI and I tried Windows_Update_SwarmUI to update it to v81 released yesterday, but it is not being updated to v81. Do I need to reinstall swarmUI again? :(
LEEJIYULL
When will wan2.2 animate be available in swarmui?
Furkan Gözükara
just checked i dont see error. give full message here
Chris
Hi Furkan, model downloader app seems broken, authentification or certificate invalid for downloading?
Furkan Gözükara
v52 installer fixes this with pip install soxr==0.5.0.post1 . what you mean by you updated?
Nicolas Giarrusso
Hi!! After installing ComfyUI and SwarmUI on Massed Compute, I get this error when trying to use the MultiTalk workflows. I update the nodes, but this log appears and I can't open ComfyUI. [ReActor] - STATUS - Running v0.6.1 in ComfyUI Torch version: 2.8.0+cu129 ComfyUI-GGUF: Allowing full torch compile Critical nanobind error: refusing to add duplicate key "SOXR_FLOAT32_I" to enumeration "soxr.soxr_ext.soxr_datatype_t"! Aborted (core dumped) (venv) Ubuntu@0151-dsm-prxmx30182:~/Downloads/Comfy_UI_V52/ComfyUI$
Furkan Gözükara
i dont know WAS Node Suite. what is it for?
Taiga
WAS Node Suite breaks your installed comfiUI. I have to be able to batch generate images and 'save image' works very well for that.
Furkan Gözükara
use demo images and prompts to understand. sadly sage attention not working with every gpu yet
Yoni Nakache
two issues while using the latest ComfyUI V49 & SwarmUI V77 (Win11,Nvidia A6000, Driver Version: 561.17 CUDA Driver Version: 12.6 (nvitop output) 1. i change model qwen image edit to architecture "qwen image edit", and tried the 3 preset for qwen image edit from the v19 presets file, but nothing is changed in the pictures, or minimal change without direct relation to the promt 2. when using --use-sage-attention even when enabling Advanced Sampling and make Preferred DType = Default (16 bit), still getting black images, so i just removed the --use-sage-attention in the backend compyUI setting. do you have a clue on what can i check/change? swarmUI debug log: https://paste.denizenscript.com/View/135771
Furkan Gözükara
which one fails give me full details so i can test and fix
Furkan Gözükara
it is fixed like 2 days ago use latest zip file overwrite older files
cool1
When I click on the models downloader option to download Flux Krea it says "User Access Token "read_gated" is expired" and doesn't download the model.
Jay
Thanks for the new qwen edit tutorial doc! I’m noticing that right now I can’t download the lesser vram models from the model downloader, it fails. The main model in the bundle is a little much for my laptop but I can’t seem to get the other ones downloaded
Furkan Gözükara
thank you so much. preparing an epic tutorial for Qwen Edit model it is amazing
Furkan Gözükara
must be doing something wrong. recording tutorial today watch it when published
cool1
When I press the preset in SwarmUI (eg. Qwen Image Edit Fast 12 Steps), and then, with the prompt entered, click "generate" it says "Invalid value for parameter LoRAs: Invalid value for param LoRAs - 'Qwen-Image-Lightning-8steps-V1.1' - must be one of: `(None)`". I've tried updating ComfyUI with Windows_Update_ComfyUI.bat today (and restarted the server in SwarmUI) and I had tried updating it with "update all" from ComfyUI Manager V3.36 last night - though it gave errors last night. It seems like it's the Loras bit that isn't working. I can use the Qwen Image edit models without the preset but they don't seem to be working well. eg. it doesn't seem to let you just add something (like a hat) without it changing too much of the whole image. Maybe I'm not using the best settings but I tied different image creativity settings and tried changing cfg scale sometimes.
Neil Rhodes
This is not a request for help, just praise where it's due. I've asked for help many times, and in the end, we got there. You helped me swiftly and calmly, even if i was a bit frustrated with things not working. Your programs and scripts work. Ive tried similar things from Aitrepreneur and nothing but problems with those installers. I just dont get that whole Confi thing, it breaks i do the wrong thing and no idea what I'm doing to get it working again. You use SWARM, which is a little like the A1111 interface, which i started using back when A1111 was new. Again, i just want to thank you for your hard work and assistance. Truly a GAOT of Patreon help and the hours you put in must be uncountable!! Thanks again! Im looking forward to trying Qwen Image Edit!! :)
Furkan Gözükara
thank you fixed with v73 zip file and Amazing_SwarmUI_Presets_v17 file
Furkan Gözükara
i will investigate this. perhaps it can be imported into comfyui workflow and changes made there
Robert Arsene
Great Job Furkan! We appreciate the work you do. Do you think it is possible at some point to have longer videos generated automated with Wan 2.2 by using last frame in SwarmUI?
JaniS69
Qwen Image 8 Steps Ultra Fast 2x Upscale - it seems missing from the Amazing_SwarmUI_Presets_v16.json
Furkan Gözükara
are you image to video or text to video? i will look if there are better LoRAs now and update them hopefully
Furkan Gözükara
yep my very next thing
Iván López
Hi Furkan, regarding wan2.2 with RTX 3090 I am getting worst results than wan2.1. I am using the presets but for some reason I get a lot of grain and blur in dynamic cases. Any idea on what parameter I can touch to improve this? Thanks!!
Iván López
Hi Furkan! Thanks for these updates. Is there any plan to do a tutorial on training Lora with Qwen?
Furkan Gözükara
hi thanks. to update swarmui just run the Windows_Update_SwarmUI.bat file which comes with zip file
juanmiguelra
Primero felicitar el gran trabajo que realiza SEcourses , me quito el sombrero de la capacidad de este señor. Felicidades!!!!. Como hago para actualizar SwarmUI?? se que esta explicado en algún sitio, pero... no encuentro la manera. First, I'd like to congratulate SEcourses on the great work they're doing. I take my hat off to this guy's skills. Congratulations! How do I update SwarmUI? I know it's explained somewhere, but... I can't find a way.
Furkan Gözükara
what is WebAgent for? i cant make bytedance/XVerse work in swarmui but i can make gradio app
Anshul Gupta
can you make i an 1click installer for Alibaba-NLP/WebAgent "https://github.com/Alibaba-NLP/WebAgent" and also can you make it possible to run bytedance/XVerse "https://github.com/bytedance/XVerse" inside swarm ui setup with model downlload. Please do this two task
Furkan Gözükara
did you manually install comfyui and gave its backend? this looks like case of outdated comfyui
Daniels MV
Getting error with WAN 2.2 : KSAMPLER ERROR have 36 channels, but got 32. ComfyUI on runpod.
Furkan Gözükara
did you change resolution? if not your comfyui is not up to date
Glenn Grillo
getting this error on using hiend preset of wan image -- 2025-08-07 18:40:09.685 [Warning] [ComfyUI-0/STDERR] raise NotImplementedError("Got 5D input, but bilinear mode needs 4D input") 2025-08-07 18:40:09.685 [Warning] [ComfyUI-0/STDERR] NotImplementedError: Got 5D input, but bilinear mode needs 4D input 2025-08-07 18:40:09.685 [Warning] [ComfyUI-0/STDERR]
Furkan Gözükara
I don't understand what you mean. We already set folder name according to comfyui or swarm ui selection
Furkan Gözükara
we already have checkbox which it remembers. check the gui and you will see all lowercase option
Taiga
your 'download manager' let's the person set the lora folder based on the system used (comfyUI or not). Yet, your swarm setup defaults to Lora... would you please get your stuff to be consistent.
Taiga
WOULD YOU PLEASE FLATTEN THE CASE OF YOUR DEFAULT DIRECTORIES TO LOWER-CASE? Every single time I have to install your 'update of the week', I have to go through and 'fix' off of your capitalized directory names. Every other installation of comfyUI is flattened.
Furkan Gözükara
hello please follow this video step by step it just recently made and still working : https://youtu.be/8cMIwS9qo4M 21:32 Part 2: Massed Compute Cloud GPU Tutorial 22:03 Massed Compute - Deploying a GPU Instance (H100) 23:40 Massed Compute - Setting Up the ThinLinc Client & Shared Folder 25:07 Massed Compute - Connecting to the Remote Machine via ThinLinc 26:06 Massed Compute - Transferring Files to the Instance 27:04 Massed Compute - Step 1: Installing ComfyUI 27:39 Massed Compute - Step 2: Installing MultiTalk Nodes 28:11 Massed Compute - Step 3: Downloading Models with Ultra-Fast Speed 30:22 Massed Compute - Step 4: Launching ComfyUI & First Generation 32:45 Massed Compute - Accessing the Remote ComfyUI from Your Local Browser 35:07 Massed Compute - Downloading Generated Videos to Your Local Computer 36:08 Massed Compute - Advanced: Integrating with the Pre-installed SwarmUI 38:06 Massed Compute - Crucial: How to Stop Billing by Deleting the Instance
cjpo
Backend issue with Massedcompute. If I am using Massedcompte, is the sequence of things still Install latest ComfyUI --> Run SwarmUI model downloader --> Run the cloudflareSwarmUI on thinlinc desktop --> then when inside SwarmUI change the backend main.py to the main.py in the ComfyUI I installed first in the download folder where it all extracted? It doesn't seem to work though it does take a while to load the backend, it seems like it's running but then say when I go to load a model it fails to load. I just wonder if I have the sequence of things wrong?
Furkan Gözükara
It is the most powerful ui at the moment. Sadly if we want to use latest tech immediately we need it. But I will keep gradio apps too and hopefully they will get better soon
djbuzz
Thanks for your work. I don't want to use SwarmUI, though. It complicates everything and the risk of being stuck with errors is too high. Please adjust your install processes for not using swarmUI
Furkan Gözükara
you are right. shown in last video as well : https://youtu.be/8MvvuX4YPeo
Taiga
"Definitely use our ComfyUI solo backend installer otherwise some of the newest stuff may fail" - you should link these types of references.
Furkan Gözükara
you probably lacking RAM too. how much RAM you have? normally it does automatic block swapping so i would expect it to work
Brett Baker
Thank you for adding this model to SwarmUI! I've installed WAN 2.2 from the "Windows_Start_Download_Models_App.bat" file and I've imported the presets from "Amazing_SwarmUI_Presets_v7.json" When I apply the "Wan 2.2 Image to Video 8 Steps Official Workflow" preset and import a 960 x 960 image, I get an "Out of Memory" error. I have an RTX 3080 Ti with 12 gb of ram. Do I need a more powerful GPU to run this model?
Furkan Gözükara
thanks let me add into downloader GGUF too
Anthony
https://huggingface.co/ND911/flux1_krea_dev_GGUFs
Anthony
You are going to want to get a SSD or NVME
DanO..
I just put my new 20TB hard drive in so I can now play with all this. I just installed your Comfy backend and Swarm/model downloader. I can't see the new Krea version of Flux. I tried restarting. I tried updating Swarm/Model Downloader. Still no. If you could, the model loader should check a list you maintain online every time it starts so any models you add are automatically added.
Furkan Gözükara
i plan to make a tutorial for Wan VACE video to video but not even it is true video to video like FLUX Kontext editing
Furkan Gözükara
sadly i never used it so dont know how to make. did you ask to the developer in official discord channel? https://discord.gg/unTKKeKFDf
Furkan Gözükara
evet bunla ilgili bir tutorial planlıyorum iyi oldu hatırlattığın
Serkan köksal
Hacam merhaba, sage attention gibi şu nunchaku'ya da bir el atsanız harika olur. İçinde nunchaku gömülü bir comfy versiyonu oluşturabilir misiniz? Flux Context ile ilgili ne kadar workflow bulduysam hepsi nunchaku kullanıyor ve ben bir türlü kurmayı beceremedim. uygun whl dosyasını kurmama ve doğrulamama rağmen iş akışları sorun çıkarıyor. Bunu çözse çözse süpermen çözer diye size yazayım dedim :)
peter realis
Hi Furkan! after a long break i tried to produce seamless textures again. This worked wonderfully with auto1111 / SD_XL. Now I have tested with the current Swarm UI / Stable Diffusion 3.5-Large & Flux Dev / Seamless Tileable = True but the textures have unsightly edges and are therefore not seamless. Are there any other settings to consider? I work with Windows.
couturecz
is there any tutorial for video 2 video?
Furkan Gözükara
well definitely pod problem in that case. get a new pod sadly. currently we start at 7861 port so you can try proxy prots 7860 7861 and 7862. if still inaccesiable get a new pod
ChudJones
Hello Furkan, I'm following the tutorial for Runpod and hitting a problem. I cannot access SwarmUI from any port on Runpod. ComfyUI installs and runs fine. Same with the Gradio Model Downloader. But SwarmUI, when running, is not accessible on any port on my pod. I've opened ports 7860-7864 on the pod, and it runs in that range, but can't be accessed.
Furkan Gözükara
extract zip file and overwrite files. then run Windows_Update_SwarmUI.bat and it should update
Matt Lowe
What is the proper way to update a swarmUI installation, say I already installed using the v26 zip file. what do i do with the v57 file
Furkan Gözükara
i didnt test video extend yet sadly. planning to cover it later hopefully
ChudJones
Thanks for the great work on this. Is Video Extend working with FusionX image2video? I've enabled it and it doesn't seem to do anything.
Furkan Gözükara
you can see from debug logs. also check your task manager
Walker4k
3090, installed as instructed on an independent, clean hardrive, but it just seems to stall on model upload, esp Flax Dev. I'm not really a fan of Swarm UI, the logs don't show that a model is loading or at what stage it is or if there is an issue with the model
Furkan Gözükara
follow this video it shows how to install comfyui properly. that auto installs : https://youtu.be/fTzlQ0tjxj0
Furkan Gözükara
thanks fixed with v55 just updated
Furkan Gözükara
nope you can't
Furkan Gözükara
can you elaborate this more i didnt understand?
Furkan Gözükara
look all tiling on all screens and enable and use Q6 or Q5 of fusionx
Furkan Gözükara
i updated workflow thus we now use fp32 version. that is why bundle doesnt have it
Furkan Gözükara
yes it is for overwriting previous one because of swarmui model naming stuff. true it will redownload. I2V thing looks hard best to delete manually
Daniel Cardona Ramirez
Please help! I'm struggling A LOT while trying to use --use-sage-attention because the backend doesn't start!! I've tried a lot of things. One person asked you how to install triton on youtube and you only replied: "it is just activate venv and pip install pre compiled wheel" but I don't know what you mean by this nor how to do it. I'm rocking a 4090 so it should work right? I don't understand why it works on my PC with 4070 ti Super but not on the 4090 one.
Taiga
ComfyUI Folder Structure (e.g. 'loras' folder) doesn't stay checked (grrrr)
Taiga
Sorry, this might be offtopic: can you "retrain" a sd lora to be flux ? Question is a mess due to ignorance on my behalf.
Taiga
OK! downloader now saves if you select to save it. The only thing that I would change on your next release is that you 'flatten' the Capitalized folder names; I have to go through and fix them each time.
Chris
Hi Furkan, no matter what, i always get a OOM error during generation of multitalk video, even when selecting 480p 10 seconds low vram preset. I have 32GB RAM and 3060 12GB VRAM. reboot, no other windows open or other processes running. any ideas?
Chris
Hi Furkan, i think the file "WanVideo_2_1_Multitalk_14B_fp8_e4m3fn.safetensors" is missing in your downloader. 480p low vram multitalk wants to load this file, but it is not present in the bundle? i downloaded it manually from huggingface.
Chris
Hi Furkan, umt5_xxl_fp8_e4m3fn_scaled.safetensors always gets downloaded again, even if it already is present in the download / models folder. Can you check? Also, can you implement a feature to only select either I2V or T2V elevant files in the bundles? Maybe with a checkbox. So we can save space when T2V or I2V is selectively needed, but not both models. Thank you for your superb work!
Furkan Gözükara
i think safe to ignore
Taiga
don't know if it's related to V50 install: [Error] [WebAPI] Error handling API request '/API/DescribeModel' for user 'local': Model not found. 13:38:11.801 [Error] [WebAPI] Error handling API request '/API/DescribeModel' for user 'local': Model not found. if it matters, models is in d:\models everything else seems to be found.
Furkan Gözükara
very interesting. i tested different browsers and also on cloud machines all working. do you see any error on cmd? can you please report here with screenshots : https://github.com/gradio-app/gradio/issues also message me discord i will give you another bat file that will install older specific gradio lets see if that improves
Furkan Gözükara
install error. reinstall comfyui and swarmui make a fresh install use python 3.10.11
revoerom
I am also getting a blank screen for the downloader. (v50, was working 2days ago, blank screen on http://127.0.0.1:7860). tried different browser/private/different installation on different drive. any suggestions?
Taiga
No backends match the settings of the request given! Backends refused for the following reason(s): - Request requires flag 'frameinterps' which is not present on the backend
Furkan Gözükara
nope i have no issues. try with different browser and private window
Chris
Hi Furkan, v50 Model Downloader only loads a white empty page? Can you reproduce?
Furkan Gözükara
can you try v50? i used so many places all worked
Taiga
The model downloader v49 doesn't keep the 'base download path'. it always resets to swarmui path.
Furkan Gözükara
this video explains logic i am same : https://youtu.be/fTzlQ0tjxj0
Taiga
You have, as I do, SwarmUI in a different directory than comfyUI. How are they connected?
Furkan Gözükara
there is literal full video what error you getting? i just made fresh install of comfyui and swarmui it works perfect : 1st follow requirements : https://youtu.be/DrhUHnYfwC0 2nd follow this video : https://youtu.be/fTzlQ0tjxj0
Astral Unification
Hi I am having trouble getting SwarmUi to work i have install and follow all your instructions but just won't work, can i get some help please? Is there any way to send some screen shots?
Furkan Gözükara
true they overwrite each other. you can put them into separate sub folders but then you need to modify paths. it is really hard to manage all with different names though. so best is put them into sub folders and just look folder names and change them :D or install 1 by 1 so overwrite will be fine
What to watch high
Love the work, thanks! I want to suggest when you create installer zips to use unique names for the instructions. I install several of your installers on Runpos and a few of the instructions for how to launch get overwritten because they are "RunPod_Install_Instructions.txt"
Furkan Gözükara
weird you should report here : https://github.com/mcmonkeyprojects/SwarmUI/issues
Justin Kupka
When i go to the swarm UI backends page after installing it and the comfy ui installed. there's no buttons to add a new backend.
Furkan Gözükara
it shows output in preview in my latest video : https://youtu.be/adF9X9E0Chs
Chris
Hi Furkan, feature request: Do you know if it is possible that when generating a picture, that the output is shown in the preview thumbnail, before the segment:yoloface begins? so we can cancel generation when we see that the base output is bad, so generation of segment:yoloface is not neccessary and better start with new generation. would save time i think!
Furkan Gözükara
yolo was broken with torch 2.7 that is why i did. but now i updated yolo with working version. i think yolo better
Chris
Another question, i saw in your last video you did only segment:face, not segment:man_yolo.pt,threshold,strongness.... Is this better or does not matter?
Furkan Gözükara
yes i recommend FusionX and Forcing LoRA comparison for your case. may work better. These 2 are better than FramePack since these are Wan 2.1 not hunyuan
Chris
Hi Furkan, so do you prefer this (Wan + FusionX) now over CausVid? And over Framepack i guess also?
Furkan Gözükara
i need more info. you can email me logs : monstermmorpg@gmail.com - logs on cmd
Nara Dhipa
I have followed all the steps in how to install swarmUI but it won't run. It just stops for a long time
Furkan Gözükara
let me make a video today to clarify stay subscribed
Yannis
Hey something isn't clear to me. Which ports are we supposed to open on runpod for running Swarm on ComfyUI backend ? Also, after installing both properly, what is the command to start swarm each time we boot the server ? It isn't indicated in the runpod instructions txt file. Thanks !
Furkan Gözükara
main py file comes after you install comfyui : https://www.patreon.com/posts/105023709 - also if you upgrade gold tier i can connect your pc and install
Javi Empredu
I cannot find the main.py file.... followed all the video step by step and installed python, and all recommended programs... so frustrated...
Furkan Gözükara
very hard to know. you can get debug logs and give me link (pastebin) so maybe there is an error. if you upgrade gold tier i can connect your pc and check out
Brett Baker
Thanks so much for this update! I followed your steps, updated SwarmUI and got no errors. I followed your steps but now all of my finished videos are a blurry mess. Any thoughts as to why this is happening?
Furkan Gözükara
great ty
Furkan Gözükara
hi please try again with latest version. path like this should work : C:\teacache tutorial\thumbs
Taiga
should I set the path to models with forward slashes instead of backslashes? Any model load falis. [Error] Error loading model 'D:/FLUX_V3/ComfyUI_windows_portable/ComfyUI/models/diffusion_models/hidream_i1_dev_fp8.safetensors' (arch=hidream-i1) on backend 0 (ComfyUI Self-Starting): System.Net.WebSockets.WebSocketException (0x80004005): The remote party closed the WebSocket connection without completing the close handshake. ---> System.IO.IOException: Unable to read data from the transport connection: An existing connection was forcibly closed by the remote host.. ---> System.Net.Sockets.SocketException (10054): An existing connection was forcibly closed by the remote host. --- End of inner exception stack trace --- at System.Net.Sockets.Socket.AwaitableSocketAsyncEventArgs.ThrowException(SocketError error, CancellationToken cancellationToken) at System.Net.Sockets.Socket.AwaitableSocketAsyncEventArgs.System.Threading.Tasks.Sources.IValueTaskSource.GetResult(Int16 token) at System.Net.Http.HttpConnection.ReadBufferedAsyncCore(Memory`1 destination) at System.Runtime.CompilerServices.PoolingAsyncValueTaskMethodBuilder`1.StateMachineBox`1.System.Threading.Tasks.Sources.IValueTaskSource.GetResult(Int16 token) at System.Net.Http.HttpConnection.RawConnectionStream.ReadAsync(Memory`1 buffer, CancellationToken cancellationToken) at System.Runtime.CompilerServices.PoolingAsyncValueTaskMethodBuilder`1.StateMachineBox`1.System.Threading.Tasks.Sources.IValueTaskSource.GetResult(Int16 token) at System.IO.Stream.ReadAtLeastAsyncCore(Memory`1 buffer, Int32 minimumBytes, Boolean throwOnEndOfStream, CancellationToken cancellationToken) at System.Runtime.CompilerServices.PoolingAsyncValueTaskMethodBuilder`1.StateMachineBox`1.System.Threading.Tasks.Sources.IValueTaskSource.GetResult(Int16 token) at System.Net.WebSockets.ManagedWebSocket.EnsureBufferContainsAsync(Int32 minimumRequiredBytes, CancellationToken cancellationToken) at System.Runtime.CompilerServices.PoolingAsyncValueTaskMethodBuilder`1.StateMachineBox`1.System.Threading.Tasks.Sources.IValueTaskSource.GetResult(Int16 token) at System.Net.WebSockets.ManagedWebSocket.ReceiveAsyncPrivate[TResult](Memory`1 payloadBuffer, CancellationToken cancellationToken) at System.Net.WebSockets.ManagedWebSocket.ReceiveAsyncPrivate[TResult](Memory`1 payloadBuffer, CancellationToken cancellationToken) at System.Runtime.CompilerServices.PoolingAsyncValueTaskMethodBuilder`1.StateMachineBox`1.System.Threading.Tasks.Sources.IValueTaskSource.GetResult(Int16 token) at System.Threading.Tasks.ValueTask`1.ValueTaskSourceAsTask.<>c.<.cctor>b__4_0(Object state) --- End of stack trace from previous location --- at SwarmUI.Utils.Utilities.ReceiveData(WebSocket socket, Int32 maxBytes, CancellationToken limit) in D:\SwarmUI_Model_Downloader_v40\SwarmUI\src\Utils\Utilities.cs:line 290 at SwarmUI.Builtin_ComfyUIBackend.ComfyUIAPIAbstractBackend.AwaitJobLive(String workflow, String batchId, Action`1 takeOutput, T2IParamInput user_input, CancellationToken interrupt) in D:\SwarmUI_Model_Downloader_v40\SwarmUI\src\BuiltinExtensions\ComfyUIBackend\ComfyUIAPIAbstractBackend.cs:line 325 at SwarmUI.Builtin_ComfyUIBackend.ComfyUIAPIAbstractBackend.LoadModel(T2IModel model, T2IParamInput upstreamInput) in D:\SwarmUI_Model_Downloader_v40\SwarmUI\src\BuiltinExtensions\ComfyUIBackend\ComfyUIAPIAbstractBackend.cs:line 955 at SwarmUI.Backends.BackendHandler.LoadModelOnAll(T2IModel model, Func`2 filter) in D:\SwarmUI_Model_Downloader_v40\SwarmUI\src\Backends\BackendHandler.cs:line 691 02:05:48.316 [Warning] Tried 1 backends but none were able to load model 'hidream_i1_dev_fp8.safetensors'
Ec Jep
Question: following your steps it will generate a strange distorted video and then any subsequent tries it gives me this error "ComfyUI execution error: The size of tensor a (1600) must match the size of tensor b (19968) at non-singleton dimension 1". The only way to clear it is to restart swarmui. I have updated to the latest version. update and solution: the preset wasn't changing "init image creativity" to 0. Once I fixed that it works as normal. Thx for your latest updates and video explanations!
Furkan Gözükara
hi can you give me their repo links?
Kahlid Seid
Hi, can you please add vast-Ai holo part Hi3DGen image to 3D
Furkan Gözükara
only if you want to download models into your comfyui. i have shown in this video : https://youtu.be/XNcn845UXdw
Tenally
Hi, thanks for the installers. Things seem to have installed corrected without issue. On the SwarmUI Model Downloader, should I check the "ComfyUI Folder Structure (e.g. 'loras' folder) , next to the "Base Download Path (SwarmUI/Models)"? I could not find that in the video. Thank you!
stphnvdb321
Sorry, I should have watched your video before. I only used your SwarmUI installer before. Now it's all working, even triton is installed :). But still no Rife in the video and image2video sttings :(
stphnvdb321
Me again. I just made a new installation of SwarmUI. I used your installer Windows_Install_SwarmUI.bat. I have no Rife and Sage attention still doesn't wok. It looks like I have a "regular" SwarmUI, not your modified version. What did I wrong? Any idea? Thank you. CausVid is great!
Furkan Gözükara
you need to install git. follow requirements tutorial : https://youtu.be/DrhUHnYfwC0
Furkan Gözükara
i will also cover that soon hopefully making a new video now, under 1 minute 2 second wan 2.1 videos
Iván López
Hello Mr Furkan! Is this supporting also Wan2.1 VACE (distilled) I see it is very powerful as reference video generation. Thanks!!
Taiga
on initial windows_start_swarmui.bat, I see this error during startup: [Error] Failed to get git commit date: System.ComponentModel.Win32Exception (2): An error occurred trying to start process 'git' with working directory 'D:\SwarmUI_Model_Downloader_v35\SwarmUI'. The system cannot find the file specified. at System.Diagnostics.Process.StartWithCreateProcess(ProcessStartInfo startInfo) at System.Diagnostics.Process.Start(ProcessStartInfo startInfo) at SwarmUI.Utils.Utilities.RunGitProcess(String args, String dir, Boolean canRetry) in D:\SwarmUI_Model_Downloader_v35\SwarmUI\src\Utils\Utilities.cs:line 1140 at SwarmUI.Core.Program.<>c.<b__24_4>d.MoveNext() in D:\SwarmUI_Model_Downloader_v35\SwarmUI\src\Core\Program.cs:line 194 00:54:56.490 [Error] Internal error in async task: System.ComponentModel.Win32Exception (2): An error occurred trying to start process 'git' with working directory 'D:\SwarmUI_Model_Downloader_v35\SwarmUI'. The system cannot find the file specified. at System.Diagnostics.Process.StartWithCreateProcess(ProcessStartInfo startInfo) at System.Diagnostics.Process.Start(ProcessStartInfo startInfo) at SwarmUI.Utils.Utilities.RunGitProcess(String args, String dir, Boolean canRetry) in D:\SwarmUI_Model_Downloader_v35\SwarmUI\src\Utils\Utilities.cs:line 1140 at SwarmUI.Core.Program.<>c.< b__24_3>d.MoveNext() in D:\SwarmUI_Model_Downloader_v35\SwarmUI\src\Core\Program.cs:line 172 --- End of stack trace from previous location --- at SwarmUI.Utils.Utilities.<>c__DisplayClass53_0.<b__0>d.MoveNext() in D:\SwarmUI_Model_Downloader_v35\SwarmUI\src\Utils\Utilities.cs:line 623
Furkan Gözükara
the model downloader will download them use it easiest way
Neil Rhodes
Upgraded to 5090 and redownloading this does the FLUX option add Flux fill model? will i need to find these models again and install manually? i forgot to save them
Furkan Gözükara
my installer auto installs that. did you use it? make a fresh install and send me logs : monstermmorpg@gmail.com. also dont use python 3.13 wont work.
stphnvdb321
I'm getting this error if I want to install Sage Attention: "To use the `--use-sage-attention` feature, the `sageattention` package must be installed first." How to do it? Thanks.
Furkan Gözükara
well i may make but i am working on video upscaler right now.
stphnvdb321
Thanks for adding LTXV 13B support! So let's install SwarmUI. I've never used it before (but you have a long video ;). I think it wasn't possible to create an LTXV installer for Gradio?
Furkan Gözükara
yes it should work as a swarmui backend
Larry Boles
Any Comfy only variations of these workflows? I'd love to give a go but don't want to install yet another frontend.
Pirotech
i got this error when installing ComFYUI
Pirotech
2025-05-08 12:03:00.690 [Debug] [ComfyUI-0/STDERR] Checkpoint files will always be loaded safely. 2025-05-08 12:03:00.736 [Warning] [ComfyUI-0/STDERR] Traceback (most recent call last): 2025-05-08 12:03:00.737 [Warning] [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\main.py", line 137, in 2025-05-08 12:03:00.737 [Warning] [ComfyUI-0/STDERR] import execution 2025-05-08 12:03:00.738 [Warning] [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\execution.py", line 13, in 2025-05-08 12:03:00.738 [Warning] [ComfyUI-0/STDERR] import nodes 2025-05-08 12:03:00.739 [Warning] [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\nodes.py", line 22, in 2025-05-08 12:03:00.739 [Warning] [ComfyUI-0/STDERR] import comfy.diffusers_load 2025-05-08 12:03:00.739 [Warning] [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\comfy\diffusers_load.py", line 3, in 2025-05-08 12:03:00.740 [Warning] [ComfyUI-0/STDERR] import comfy.sd 2025-05-08 12:03:00.740 [Warning] [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\comfy\sd.py", line 7, in 2025-05-08 12:03:00.741 [Warning] [ComfyUI-0/STDERR] from comfy import model_management 2025-05-08 12:03:00.741 [Warning] [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\comfy\model_management.py", line 221, in 2025-05-08 12:03:00.742 [Warning] [ComfyUI-0/STDERR] total_vram = get_total_memory(get_torch_device()) / (1024 * 1024) 2025-05-08 12:03:00.742 [Warning] [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\comfy\model_management.py", line 172, in get_torch_device 2025-05-08 12:03:00.742 [Warning] [ComfyUI-0/STDERR] return torch.device(torch.cuda.current_device()) 2025-05-08 12:03:00.743 [Warning] [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\venv\lib\site-packages\torch\cuda\__init__.py", line 1026, in current_device 2025-05-08 12:03:00.743 [Warning] [ComfyUI-0/STDERR] _lazy_init() 2025-05-08 12:03:00.744 [Warning] [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\venv\lib\site-packages\torch\cuda\__init__.py", line 363, in _lazy_init 2025-05-08 12:03:00.745 [Warning] [ComfyUI-0/STDERR] raise AssertionError("Torch not compiled with CUDA enabled") 2025-05-08 12:03:00.745 [Warning] [ComfyUI-0/STDERR] AssertionError: Torch not compiled with CUDA enabled 2025-05-08 12:03:01.029 [Info] Self-Start ComfyUI-0 unexpectedly exited (if something failed, change setting `LogLevel` to `Debug` to see why!) 2025-05-08 12:03:01.031 [Debug] Status of ComfyUI-0 after process end is LOADING 2025-05-08 12:03:01.031 [Info] Self-Start ComfyUI-0 had errors before shutdown: [ComfyUI-0/STDERR] Adding extra search path checkpoints C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\Stable-Diffusion [ComfyUI-0/STDERR] Adding extra search path vae C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\VAE [ComfyUI-0/STDERR] Adding extra search path loras C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\Lora [ComfyUI-0/STDERR] Adding extra search path loras C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\LyCORIS [ComfyUI-0/STDERR] Adding extra search path upscale_models C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\ESRGAN [ComfyUI-0/STDERR] Adding extra search path upscale_models C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\RealESRGAN [ComfyUI-0/STDERR] Adding extra search path upscale_models C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\SwinIR [ComfyUI-0/STDERR] Adding extra search path upscale_models C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\upscale-models [ComfyUI-0/STDERR] Adding extra search path upscale_models C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\upscale_models [ComfyUI-0/STDERR] Adding extra search path embeddings C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\Embeddings [ComfyUI-0/STDERR] Adding extra search path hypernetworks C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\hypernetworks [ComfyUI-0/STDERR] Adding extra search path controlnet C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\controlnet [ComfyUI-0/STDERR] Adding extra search path clip C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\clip [ComfyUI-0/STDERR] Adding extra search path clip_vision C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\clip_vision [ComfyUI-0/STDERR] Adding extra search path unet C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\unet [ComfyUI-0/STDERR] Adding extra search path diffusion_models C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\diffusion_models [ComfyUI-0/STDERR] Adding extra search path gligen C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\gligen [ComfyUI-0/STDERR] Adding extra search path ipadapter C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\ipadapter [ComfyUI-0/STDERR] Adding extra search path yolov8 C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\yolov8 [ComfyUI-0/STDERR] Adding extra search path tensorrt C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\tensorrt [ComfyUI-0/STDERR] Adding extra search path clipseg C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\clipseg [ComfyUI-0/STDERR] Adding extra search path style_models C:\SwarmUI_Model_Downloader_v31\SwarmUI\Models\style_models [ComfyUI-0/STDERR] Adding extra search path custom_nodes C:\SwarmUI_Model_Downloader_v31\SwarmUI\src\BuiltinExtensions\ComfyUIBackend\DLNodes [ComfyUI-0/STDERR] Adding extra search path custom_nodes C:\SwarmUI_Model_Downloader_v31\SwarmUI\src\BuiltinExtensions\ComfyUIBackend\ExtraNodes [ComfyUI-0/STDERR] Checkpoint files will always be loaded safely. [ComfyUI-0/STDERR] Traceback (most recent call last): [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\main.py", line 137, in [ComfyUI-0/STDERR] import execution [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\execution.py", line 13, in [ComfyUI-0/STDERR] import nodes [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\nodes.py", line 22, in [ComfyUI-0/STDERR] import comfy.diffusers_load [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\comfy\diffusers_load.py", line 3, in [ComfyUI-0/STDERR] import comfy.sd [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\comfy\sd.py", line 7, in [ComfyUI-0/STDERR] from comfy import model_management [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\comfy\model_management.py", line 221, in [ComfyUI-0/STDERR] total_vram = get_total_memory(get_torch_device()) / (1024 * 1024) [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\comfy\model_management.py", line 172, in get_torch_device [ComfyUI-0/STDERR] return torch.device(torch.cuda.current_device()) [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\venv\lib\site-packages\torch\cuda\__init__.py", line 1026, in current_device [ComfyUI-0/STDERR] _lazy_init() [ComfyUI-0/STDERR] File "C:\Comfy_UI_V27\ComfyUI\venv\lib\site-packages\torch\cuda\__init__.py", line 363, in _lazy_init [ComfyUI-0/STDERR] raise AssertionError("Torch not compiled with CUDA enabled") [ComfyUI-0/STDERR] AssertionError: Torch not compiled with CUDA enabled 2025-05-08 12:03:01.031 [Error] Self-Start ComfyUI-0 on port 7822 failed. AutoRestart ignored as this was an initial launch failure. 2025-05-08 12:03:01.432 [Debug] [Load ComfyUI Self-Starting #0] ComfyUI-0 self-start port 7822 loop ending (failed?)
Furkan Gözükara
please use v26. it auto sets a hugging face token of mine. it is not 100% complete yet but it has so many new models and let me know
Richard Postieror
Having some issues with the installer. was getting "Python was not found; run argument etc" Got it fixed, now when i try to install anything from V25 i get "Cannot access gated repo for url https://huggingface.co/black-forest-labs/FLUX.1-schnell/resolve/741f7c3ce8b383c54771c7003378a50191e9efe9/ae.safetensors. Access to model black-forest-labs/FLUX.1-schnell is restricted. You must have access to it and be authenticated to access it. Please log in." i logged facehug, and requested access. i got it granted but it still is not accessing [edit] deleted Python, maybe i installed it incorrectly? sorry, complete novice here. i just like the pretty pictures lol [/edit]
Furkan Gözükara
awesome. yes it speeds up. i will make a tutorial hopefully
Chris
Hi Furkan, maybe interesting: I copied over your sage-attention compiled comfyui backend from comfyui patreon post to swarmui directory, overwriting the comfyui part of swarmui. doing so, and adding --use-sage-attention in the swarmui server backend gives a very nice speed boost, from about 6.8sec/it to 4.2sec/it. teacache module seems to be broken with this copy method though.
Furkan Gözükara
thanks for info
Chris
Hi Furkan and all members, i want to share my success with triton installation and very good speed increase. Here is what i had to to: modify SwarmUI_Upgrade_Torch_Triton_Flash_RTX5000_Series.bat and edit python.exe pip install -U --pre triton-windows python.exe pip install huggingface_hub ipywidgets hf_transfer python.exe pip install moviepy to python.exe -m pip install........ Then in SwarmUI change prefeded dtype to FP8e5m2 (at least for my 3060 12GB) The speed increases from about 6.80s/it to 5.10s/it for my config.
Furkan Gözükara
sorry i dont have ubuntu to test there :D but you can look windows installer and perhaps apply to there
Baekdoosixt
Hello mate , thank you for this last release . It runs like a charm on windows on my 5090. But I'm running out of patience to make it run on Ubuntu . Several clean install later , i 'm unable to figure how can i make SwarmUI run on my 5090 ( i was able to make it run flawlessly on my 4090 on this same install)... if you know how can i do plesa give me a hint
Furkan Gözükara
Yes you can install. I have RTX 3090 works on them as well.
Brandon
Mostly update for 5000 series? If you already have your last version for a 3090 user should install deepspeed and the other updates
Furkan Gözükara
yep
Brett Baker
Thank you! I know the model just came out but it looks really interesting, especially for lower GPU's.
Furkan Gözükara
yes i will add hopefully asap and i will make interface better too
Brett Baker
Is it possible to add the latest video model "hunyuan_video_I2V_fp8_e4m3fn" to this SwarmUI build?
Furkan Gözükara
thanks will do :D
Hockey
Change 2024 to 2025 for the recent posts. =)
Furkan Gözükara
you are welcome. great it works.
Brett Baker
It was user error. I simply needed to choose the model and the error went away. For some reason I thought it would load the model automatically. Thank you for your quick response!
Furkan Gözükara
please show screenshot so i can comment better. you can ping me in discord
Brett Baker
I just installed the latest build. Despite downloading the necessary models, I'm still getting the error: "Model Not Found" when I launch the program. What model would it be looking for?
Furkan Gözükara
these are only text to video. hunyuan best atm for text to video. also hunyuan probably slower i didnt compare with cosmos. wait my tutorial. works perfect on rtx 3090 and i will show in tutorial hopefully
cool1
it says "Hunyuan text-to-video Models" is that also image 2 video? I've got cosmos image 2 video installed in comfyui. Does Hunyuan work better than that (if it can do image 2 video too?) on PC with 1 RTX3090 GPU which has 24 GB of VRAM (and plenty of RAM)? Does it create better quality videos and/or faster ones than cosmos image 2 video (when used with RTX3090)?
Chris
Tried them, they work better than the standard models for me!
Furkan Gözükara
yes please download v20 i noticed and fixed it :)
Nedo Braun
trying to upgrade Torch RTX 5090.bat and getting this error: D:\Ai\SECourses\SwarmUI-Sec\SwarmUI_Model_Downloader_v19>python Torch_Upgrade.py python: can't open file 'D:\\Ai\\SECourses\\SwarmUI-Sec\\SwarmUI_Model_Downloader_v19\\Torch_Upgrade.py': [Errno 2] No such file or directory
Furkan Gözükara
true
Furkan Gözükara
great
Baekdoosixt
That worked , tks a lot ;-)
Markus
I would say that it depends on what you are trying to achieve. Comfy is obviously more flexible, thus you can make more fancy stuff. I use both depending on what I need
Renzo Canepa Garay
Hey Chris! Have you tried those new Yolo models on SwarmUI?
Furkan Gözükara
pip install huggingface_hub hf_transfer
Furkan Gözükara
i tested on massed compute and runpod both works. they have ubuntu 22. what is your exact error? install huggingface_hub and hf_transfer manually it should fix
Baekdoosixt
Hello do you know how to bypass the PEP 668 "pip issue" on ubuntu ? Cause it fails to download after the choices
Furkan Gözükara
hi not that i know sadly. can you ask in their discord? i sometimes do that with prompt SR feature of grid :D https://discord.gg/KfSd3ABg
Umut Hasanoglu
Hi. Is it possible to run a prompt list in SwarmUI like on the forge and a1111 "prompt from list or text file". I use wildcards instead but it's random and I need to run the prompts in order.
Furkan Gözükara
i use face_yolov9c.pt it works best.
Philipp Capetian
Are there new prompts for the yolo models? Segment face woman does not seem to work...
Furkan Gözükara
put your image into a bigger transparent canvas png and upload that way. that way you can expand all directions
Furkan Gözükara
hi i didnt test them but if they are working better for you, you can surely use.
Chris
Hi Furkan, there are new (? new file names ?) yolo models at: https://huggingface.co/Anzhc/Anzhcs_YOLOs. are yours in the downloader still the recommended ones?
Furkan Gözükara
i would guess yes but i dont know how to. you can ask to the developer. also you can start swarmui on cloudflared on any platform
Louis L'herrou
Hello ! Is there a way to use Swarm via API on my cloud instance, would be ideal to have swarmui and all the process in the Kohya Flux fine tunning tutorial available via api call :)
Michael Liu
Yes, and how do you resize the yellow dotted border like in your config image, to outpaint top bottom left and right at the same time?
Furkan Gözükara
hi did you see latest config image file i used? https://cdn-uploads.huggingface.co/production/uploads/6345bd89fe134dfd7a0dba40/ZKIQZ90IIohFs6nwkFbEP.jpeg
Furkan Gözükara
it is hard to use i pefer swarmui
Nabil Boulezaz
Why you dont use comfyui isn't better then swarmui and forge
Michael Liu
hi furkan i tried your latest config for Flux outpainting in swarmui. however it generates black image unless i set it to Euler+Beta. it does work, but the seamings are very prominent. am i missing something?
Furkan Gözükara
Hehe yes i have optimized it to maximum :D
Error_404_unknown
WOW this is awesome!!!, I don't understand these download speeds! I'm getting 15MB/s the most I've ever seen from my ISP is 7MB/s, what magic did you do to double the speed? Thanks so much
Mark Sutherland
Thank you for that, very much appreciated. I will watch the videos and use them as a guide and see the process. Thanks again for taking the time to make the videos for people to watch and for letting me know, have a great day, cheers
Furkan Gözükara
nice info thanks
Christian Schülling
Maybe if you dont mind that the video is in german (I think you can easily use subtitles..) then maybe my videos will help you. https://www.youtube.com/watch?v=zn2H4_xOzgQ https://www.youtube.com/watch?v=1bk2IIsSMbA&t=793s https://www.youtube.com/watch?v=s8Ys_otvKBs&t=4s In the videos I show how I cloned my singing and my talking voice. So I can use it with RVC and also with TTS. I think RVC is the one you maybe forgot, it doesnt need that much audio. So please keep in mind, the videos are a few weeks old which seems ages nowadays...so as far as I know there is a much easier installation process now, because you now can use Pinokio. The RVC WebUI is now in there and you can easily install it with one click without any hustle :-) I trained my models on the audio of a part of my YouTube Clips and the singing model actually from Voice Recordings I made for my band in my homestudio. But in case you dont have enough audio stuff from yourself (what's the problem in the most cases...) you can try to use your spoken WhatsApp Messages. If you use WhatsApp Desktop you can search for the files on your hard drive (I forgot where I found them), most of them are saved somewhere on your drive and you can use it. You maybe need to convert it and you ALWAYS should hear you Training Data before training to make sure there is no unwanted stuff trained with your voice. As you can see in my videos it only took a few hours (I think maybe around 2 dependend on your Specs..) to train a model from your voice. And then the only thing you need is a source audio to replace the voice with yours. I got some really really good results and I also used my trained model later in the studio and used it to make some background choir. Surely I sung the parts before and then replaced it with AI (So keep in mind: You need source audio to convert), but I was really really useful because I could pitch the model slightly without making it sound like a chipmunk and so my AI background voices were able to sing some parts higher than I was able myself. But there are many many other use cases. I also made a good night story with RVC and TTS, it sound nearly like me but it sounds as I had a good portion of speaking training...like a radio moderator or so. I wouldnt talk like this, but as a model for Podcasts and Narration its perfect.
Furkan Gözükara
awesome
Chris
Yes, go to folder dlbackend, comfyui, there are command files to tun in order to update the underlying files, i.e. ultralyrics will be updated to 8.3. this fixed the problem for me.
Ash
Were you able to resolve this? I'm getting the same error with the newer yolo in swarmui. Previous one works just fine
Chris
Okay, will do
Furkan Gözükara
please use discord and show screenshot
Chris
I get this error in SwarmUI :(
Furkan Gözükara
thank you for support as well
Mark Sutherland
Fantastic, much appreciate both for this and all your work, thank you
Furkan Gözükara
yes i plan hopefully the best one soon
Mark Sutherland
Thank you, would it be possible for you to create a one click installer for a voice cloning software. I have heard there is one which only requires a small clip of the voice but I've forgotten the name of it, thank you
Umut Hasanoglu
thank you
Furkan Gözükara
you can download separate from here but you need to download vae as well dont forget : https://huggingface.co/Comfy-Org/mochi_preview_repackaged/tree/main/split_files
Furkan Gözükara
yes fixed it ty
Furkan Gözükara
i dont know comfyui sadly i only used with swarmui. by fail i mean it couldnt detect face accurately
Chris
Do you mean this error: ComfyUI execution error: 'Segment' object has no attribute 'detect'. I have this error with the 2 new files on many faces, even when they are prominent in the picture.
Furkan Gözükara
they work better but on some occasion they fails
Chris
Hi Furkan, are the new man_face.pt and woman_face.pt different from the file face_yolov9c.pt that you used before? Are they better or why are there now 2 files? Are they newer files / models?
Umut Hasanoglu
Also, it would be better if there were two separate selections for fp16 and fp8. It downloads both of the video models and t5 models
Umut Hasanoglu
Hi. Windows download batfile has a wrong title on number 12. It repeated the title of 11 instead of saying Mochi 1 Video Model. But it downloads the Mochi, so only the bat file is incorrect.
Furkan Gözükara
ye it works in SwarmUI. cant tell how to make it comfyui sadly since i don't know either
David Fernández
That segmentation file does not work in comfy, even with ultralitycs updated to 8.3