SwarmUI Latest Updates

SwarmUI Latest Updates

18 47 76
calendar_today agoschedule17 min read

By SECourses: FLUX, Tutorials, Guides, Resources, Training, Scripts | Original Patreon post

Get the SwarmUI Installer and Model Downloader app and the presets from here : https://www.patreon.com/SECourses/posts/swarmui-auto-and-114517862

28 August 2026 Update V172

  • Reference tokens are more robust now.
    • Variants such as @IMAGE1, @ image # 2, and are recognized and normalized automatically.
  • Prompt-label handling and makes token estimates more accurate for multi-frame continuation and Image-to-Video workflows now.
  • Better Video Continuation
    • Init Video Continuation 1.3.1 now supports 1, 5, 22, 39, or 56 context frames.
    • One-frame mode preserves the previous behavior. Multi-frame mode uses MiniMax H3's native clip conditioning for smoother transitions.
    • Replayed context frames and matching audio are removed automatically during merging, preventing duplicated footage or sound.
    • Source audio is preserved and generated audio is appended at the continuation boundary.
  • Start Frame / Init Video now accepts an image or video. Uploaded videos can be continued and optionally merged with the newly generated segment.
  • Continuation no longer consumes a Ref2VA reference-image slot.
  • MiniMax H3 Fixes
    • Fixed text-only Audio Only workflows failing to recognize SwarmUI's H3 latent canvas.
    • Fixed pasted multiline wildcard text being collapsed.
  • WD14 Tagger Hardening
    • All 19 supported model repositories are now allowlisted and pinned to reviewed, immutable Hugging Face revisions.
    • Downloaded Taggerine inference code must pass SHA-256 verification before execution.
    • Added stricter validation for model IDs, thresholds, image types and sizes, model directories, and temporary output paths.
    • Runtime package installation was removed. Missing dependencies now produce clear installation instructions instead of modifying the ComfyUI environment during generation.
    • Added dedicated security-boundary tests.
    • Install from this repo manually https://github.com/FurkanGozukara/SwarmUI-WD14Tagger
  • Automatic Installers / Updaters Reliability
    • Aggressive Git updates now clean untracked leftovers after repository resets, improving recovery from blocked or inconsistent node updates.
  • Get latest zip file extract and overwrite all files
    • Import latest presets
    • Run Windows_Install_SwarmUI.bat to update latest
  • Get latest ComfyUI and update it to latest as well
  • More latent frames will increase vram usage a lot so be careful
  • This feature added to ComfyUI presets as well read changelogs on ComfyUI

SwarmUI update screenshot 1

22 August 2026 Update V170

  • This is a very big update : LTX 2.5 video generation presets + LTX 2.5 Video Core Bundle, new FLUX 2 Klein INT8 / INT4 ConvRot HQ models, MiniMax H3 Init Audio (make the video follow any soundtrack with lipsync), live prompt token meter, Face Inpainting upgrades, a big model downloader UI upgrade and a huge installer / updater overhaul so please read all
  • LTX 2.5 video generation presets finally added : Amazing_SwarmUI_Presets_v70.json brings 6 new presets : LTX25 Text To Video - Core 2-Stage 8+3, LTX25 Image To Video - Core 2-Stage 8+3, LTX25 First Last Frame - Core 2-Stage 8+3, LTX25 Audio Reference To Video - Core 2-Stage 8+3, LTX25 Text To Video - Dev HQ 30+6 and LTX25 Image To Video - Dev HQ 30+6 (all 260822)
  • Audio Reference To Video : add a Prompt Audio file and describe the matching visuals - the native LTX 2.5 audio branch makes the video follow your audio (lipsync, timing) - Image To Video and First Last Frame presets : add an Init Image (and a Video End Frame) and describe only the motion, camera and sound instead of re-describing the still

SwarmUI update screenshot 2

  • All 6 presets are two-stage : the first stage generates at 960x544, then SwarmUI's Refine / Upscale applies the official LTX 2.5 latent spatial upscaler x2 and refines to model-aligned 1080p (1920x1088) with synchronized audio - 97 frames at 24 FPS (about 4 seconds) and H264 MP4 output by default
  • Core presets use my LTX 2.5 Distilled INT8 ConvRot Premium model with 8 + 3 steps at CFG 1 and Euler Ancestral so they are fast - Dev HQ presets use my LTX 2.5 Dev INT8 ConvRot Premium model with 30 + 6 steps at CFG 3 and Euler for maximum prompt adherence and detail
  • Everything is exposed as normal SwarmUI parameters so you can change resolution, duration, steps and the refiner settings as you like - increase Text2Video Frames for longer clips
  • Import Amazing_SwarmUI_Presets_v70.json or run Windows_Preset_Delete_Import.bat to get them - the Which_Bundles_Downloads_Which_Preset_Models.html report is regenerated for the new presets as well

SwarmUI update screenshot 3

  • LTX 2.5 Video Core Bundle added to the model downloader : 1 click downloads everything the LTX 2.5 presets need - 10 files, 64.64 GB : Dev and Distilled INT8 ConvRot Premium transformers, Gemma 4 12B INT8 ConvRot v2 text encoder (SwarmUI's current default), Gemma 4 E2B INT8 ConvRot prompt enhancer, the convolutional and the diffusion video VAE, the audio VAE, the latent spatial upscaler x2, the pixel spatial upscaler x2 IC-LoRA and the motion track control IC-LoRA
  • This one bundle covers every model referenced by all of our LTX 2.5 presets in both SwarmUI and ComfyUI, so you do not need to hunt for a second bundle to make any of them run
  • The text encoder and both VAEs are saved with the exact LTX-2 sub folder and file names SwarmUI's own automatic downloader expects, so SwarmUI never downloads a second copy - in ComfyUI mode they go to models/text_encoders/LTX-2 and models/vae/LTX-2
  • No more gated downloads for LTX 2.5 : I have mirrored the remaining official Lightricks files into my own non-gated repository, so nothing in the LTX 2.5 list asks you to log in or accept a license anymore - this covers the official Dev and Distilled Comfy INT8 ConvRot transformers, the Distilled NVFP4 transformer (for RTX 5000 Blackwell), the pixel spatial upscaler x2 IC-LoRA and the latent spatial upscaler
  • These official quantizations stay optional alternatives - my own Premium INT8 ConvRot conversions remain the default of the bundles and presets because they are higher quality, and the Gemma 4 E2B INT8 ConvRot prompt enhancer is also listed separately

SwarmUI update screenshot 4

  • Model downloader UI upgrade : every model list and every bundle Includes list is now grouped by the folder the files are saved into : Main Models, Text Encoders, VAEs, LoRAs, ControlNets, Upscalers, Vision, Detection, LLM - biggest files first, and every group header shows the file count and total GB so you instantly see what goes where

SwarmUI update screenshot 5

  • New Active Transfers panel with a real progress bar per file (percent, downloaded / total, speed, ETA) above the Activity Log - the log is chronological now with a Follow newest log entry checkbox, it refreshes twice per second, and multiple open browser tabs no longer steal each other's updates
  • The URL Downloader tab now shows its progress in the same panel instead of looking frozen, and every download reports \[STARTING\] immediately instead of showing nothing while the file metadata is fetched
  • Console progress line reordered : percent, downloaded / total size, speed and ETA come first and the long file name goes last, so a narrow CMD window never cuts off the numbers - and when the output is not a real terminal (RunPod / Massed Compute logs) progress is still printed once per second instead of being silent

SwarmUI update screenshot 6

  • New LTX2.5_Enchance_Prompt_Feed_For_LLMs.txt inside the Prompt_Generate_LTX_MiniMax_And_Presets_How_To_Use folder : an official-aligned LTX 2.5 prompting guide for LLMs - give it to ChatGPT / Claude / Gemini, tell it what you want, and it writes properly formatted LTX 2.5 prompts (text / image / audio to video, multi-shot and IC-LoRA conventions are all covered)
  • MiniMaxH3References extension 1.13.2 : videos and audio files attached with SwarmUI's own native prompt uploader are now recognized as MiniMax H3 references too - same cards, same @video1 / @audio1 tokens with autocomplete, same Ref2VA workflow
    • All extensions and custom nodes automatically installed by our installers
  • FLUX 2 Klein INT8 and INT4 ConvRot HQ models : I have compiled high quality INT8 ConvRot and INT4 ConvRot versions of FLUX 2 Klein myself - Base 9B, Distilled 9B and Distilled 9B kv - 6 new models added to the model downloader app
  • FLUX 2 Klein Core Bundle and Complete Image Generation and Editing Bundle now download the INT8 ConvRot HQ versions - almost same quality as BF16 but massively faster, same as our Krea 2 and LTX 2.5 ConvRot models

SwarmUI update screenshot 7

  • With Amazing_SwarmUI_Presets_v70.json both FLUX 2 Klein Base and FLUX 2 Klein Distilled 8 Steps presets now use the INT8 ConvRot HQ models
  • MiniMax H3 FL2V Turbo 4-Step 768p speed LoRA updated to v1.1 : the model downloader now downloads minimax_h3_fl2v_turbo_4step_v1.1_768p_bf16.safetensors instead of the old v1.0 - this is the optional 16:9 aspect ratio speed up LoRA for 768p generations from LightX2V - MiniMax-H3 Core and Low VRAM bundles download the new version automatically

SwarmUI update screenshot 8

  • MiniMax H3 Init Audio : new Init Audio parameter group added - select or upload any audio file (or even a video file, its soundtrack is used) and MiniMax H3 generates the video to follow that soundtrack exactly : lipsync, action timing, ambience - and keeps it as the exact final audio track of the output
  • Works for Text To Video, Image To Video and References To Video presets and Init Audio Match Duration makes the video length automatically follow the audio length
  • Live token meter : a real time Tokens meter now appears above the prompt and shows estimated MiniMax H3 token usage (prompt + references + resolution + duration) against the 109k budget - you instantly see when your references or duration are too big before generating
  • Face Inpainting upgrades : new Face Inpaint Faces parameter selects which detected faces get refined : 1 = biggest face (default), 2 = second biggest, 1,3 or all - faces are ranked biggest to smallest and each selected face is refined in its own pass
  • New hallucination guard prevents pasting a wrong / neighbouring face over your subject and Face Inpaint Detector can no longer block generations

SwarmUI update screenshot 9

  • Installer / Updater overhaul : Windows_Install_SwarmUI.bat and Windows_Update_SwarmUI.bat are merged into a single Windows_Install_or_Update_SwarmUI.bat
  • Windows_Start_SwarmUI.bat is now ultra fast : it starts SwarmUI directly without updating or downloading anything and rebuilds only if the code actually changed
  • Update runs are much faster : unchanged extensions (Premium Extensions, Licon MSR, Foley) are skipped with a quick revision check instead of being re-downloaded on every run - FFmpeg and Cloudflared downloads are also skipped when already latest
  • Missing Git is now auto installed (winget on Windows, apt-get on Linux) instead of failing with cryptic WinError 2 errors
  • SwarmUI is now installed as a full git clone and older shallow installs are auto completed - fixes the scary Tag list empty?! warning of SwarmUI's own update check
  • .NET 10 SDK now installs into SwarmUI/.dotnet folder, matching SwarmUI's own launcher search order and persisting on RunPod / SimplePod network volumes - DOTNET_ROOT is exported so in-app Update and Restart now rebuilds with the correct SDK and works properly
  • Linux / cloud : Cloudflared is now installed system-wide from Cloudflare's official APT repository and launching prints the public trycloudflare.com URL so you can open SwarmUI from your own computer's browser on RunPod / SimplePod / Massed Compute - new --no-cloudflared and --host options added
  • RunPod, SimplePod and Massed Compute instruction files rewritten with exact HOW TO OPEN SWARMUI steps
  • Windows_Preset_Delete_Import.bat made more robust : works from any folder and auto picks the correct Python
  • To update : get latest zip file, extract and overwrite all, then run Windows_Install_or_Update_SwarmUI.bat and import Amazing_SwarmUI_Presets_v70.json
  • Also update your ComfyUI backend to latest V127 and get the 8 new LTX 2.5 ComfyUI presets : https://www.patreon.com/SECourses/posts/comfyui-auto-2-105023709

16 August 2026 Update V168

  • MiniMax H3 bundle updated to include newest 8-steps LoRA and also necessary Yolo Face model for new automatic ComfyUI Face Inpainting feature in presets for MiniMax H3
  • SwarmUI presets are now also using new Ref2V LoRA for references presets, you can switch to FL2V LoRA and see which one performs better so easy from LoRA tab and compare

SwarmUI update screenshot 10

  • Automatic video face inpainting implemented

SwarmUI update screenshot 11

14 August 2026 Update V166

  • We have added LTX 2.5 video upscaler and MiniMax Music 3 Text to Music presets into ComfyUI : https://www.patreon.com/SECourses/posts/comfyui-auto-2-105023709
    • LTX 2.5 video upscaler preset is able to upscale any video, any resolution, any aspect ratio into 2x
    • It is a very serious generative upscale so adds lots of details
    • Therefore, try to upscale at 1 chunk to have better consistency
    • Read preset information carefully after loading into ComfyUI
    • Our SECourses Premium Upscaler Pro app also fully supports LTX 2.5 now : https://www.patreon.com/SECourses/posts/secourses-pro-150202809
      • It is standalone does not require ComfyUI and it auto downloads models when you first time use
  • To make these presets work out of the box, 2 new ComfyUI core bundles added to the downloader app as below
  • LTX 2.5 video generation presets and bundles will be added soon hopefully but its quality lower than MiniMax H3
  • Moreover, I have converted the LTX 2.5 Int8 ConvRot myself higher quality and more accurate than officially published ones
    • Someone detected one of the model is corrupt and broken - officially released ConvRot Int8 variant
    • I converted them from BF16 versions after big research and experimentation
    • So our model downloader app will download better quality Int8 ConvRot of LTX 2.5 models
  • Our Musubi Trainer convert to quant tab also now supporting LTX 2.5 Int8 ConvRot compile : https://www.patreon.com/SECourses/posts/secourses-musubi-137551634

SwarmUI update screenshot 12

12 August 2026 Update V165

SwarmUI update screenshot 13

  • Our MiniMax H3 core bundles now auto downloads both FL2VA Turbo 8-step v1.0 and FL2VA Turbo 4-step v1.0 768p
    • Older v0.1 removed from downloads and presets so you can delete it from LoRA downloads if you did download before
    • All presets now by default uses FL2VA Turbo 8-step v1.0 but you can switch it to FL2VA Turbo 4-step v1.0 768p
      • If you switch to FL2VA Turbo 4-step v1.0 768p you may be needed to overwrite Video shift value and set to 6, SwarmUI default sets to 12
    • Sadly LoRAs only made for non-reference model atm but works on both models until they publish specific new LoRA for reference model
  • Quick set megapixels feature implemented
    • Based on selected aspect ratio, it will set resolution according to your set
  • Moreover, now it will display visualization of the output based on aspect ratio and resolution

SwarmUI update screenshot 14

  • To update get the latest zip file, extract and overwrite all, run Windows_Update_SwarmUI.bat and also use Windows_Preset_Delete_Import.bat to get latest updated presets
  • Moreover get latest ComfyUI backend zip file and also update it through installer bat file : https://www.patreon.com/SECourses/posts/comfyui-auto-2-105023709

10 August 2026 Update V162

ComfyUI infinite video with MiniMax H3 tutorial published : https://youtu.be/1580ZDX-60Q

Hopefully a tutorial for SwarmUI coming soon

  • Completely Rebuilt Download Engine
    • Exact byte-range resume after cancellation, connection loss, or application restart.
    • Previously downloaded ranges are preserved instead of downloaded again.
    • One sparse staging file replaces multiple temporary chunks and the expensive final merge.
    • Significantly lower temporary disk-space requirements.
    • Models are atomically installed only after size and SHA-256 verification.
    • Corrupt or incomplete downloads can never replace a finished model.
  • Faster Hugging Face Downloads
    • Native Hugging Face Xet is now recommended and enabled automatically when available.
    • Xet automatically scales between 8 and 32 streams according to available system RAM.
    • Improved recovery from stalled transfers, rate limits, server errors, and expired signed URLs.
    • Removed the previous forced approximately 84.5 GiB Xet cache.
    • New downloads default to no chunk cache, with an optional bounded reuse mode.
  • Improved Download Controls
    • Direct URL downloads now appear in the active-download counter.
    • Cancel and Cancel All now work with direct URL downloads.
    • Completed ranges remain available after cancellation.
    • Prevented overlapping standalone URL downloads.
    • Small and large direct downloads now use the same atomic download system.
    • Improved URL validation and rejection of HTML pages masquerading as model files.
    • Safer Hugging Face and CivitAI domain detection.
  • Catalog And Interface Improvements
    • Replaced the unavailable LTX 2.3 x2 spatial upscaler v1.0 with the official v1.1 long-video hotfix : ltx-2.3-spatial-upscaler-x2-1.1.safetensors
    • Removed duplicate physical models from search and LoRA displays while keeping aliases searchable.
    • Improved automatic folder detection for filenames containing hyphens and underscores.
    • Improved mobile and narrow-screen layout.
  • Updater now detects and repairs incomplete FoleyExtension installations, including optional-image components.
  • MiniMax H3 Improvements
    • Expanded the MiniMax H3 prompt-enhancement guide to v2.3.
    • Added stable reference-roster numbering rules for multi-scene and folder-batch prompts.
    • Added clearer image, video, audio, and paired-soundtrack reference rules.
    • Corrected first-and-last-frame preset instructions to use Image To Video > Video End Image.
  • For updating please also update ComfyUI to latest as well and also this one
  • Use Windows_Update_SwarmUI.bat and import latest presets file

9 August 2026 Update V160

  • Model downloader app updated and Int4 ConvRot MiniMax H3 models bundle added as Low VRAM
    • This bundle is great for 12 GB and below GPUs
    • At the presets both ComfyUI and SwarmUI, just replace Int8 Model with below Int4 variants
    • This format Int4 ConvRot models right out of the box working with our ComfyUI (use our installers) backend and installers for SwarmUI

SwarmUI update screenshot 15

  • We are using official code of Lightricks not Kijai implementation therefore our LoRA works better
  • However, Kijai had a VRAM optimization and i just implemented it
  • You will see it like below in presets, MAX savings reduces VRAM more than 40% but may reduce quality, exact saves 15%+ but same quality
  • Make sure to use our ComfyUI installer and update it for all features to run via our bat file
  • Remember MiniMax H3 specific features will only appear when you have selected MiniMax H3 architecture model in your model selection

SwarmUI update screenshot 16

  • If you set reference image size to max like this, it improves quality and accuracy but may use more memory, default is match

SwarmUI update screenshot 17

  • The importance of Max option is very significant

SwarmUI update screenshot 18

  • Video clip references are processed differently and below explains how they processed when used as a reference

SwarmUI update screenshot 19

8 August 2026 Update v159

  • This is a very major upgrade with so many new amazing stuff so please read carefully
  • Famous Lightricks released 4 steps LoRA for MiniMax H3

SwarmUI update screenshot 20

  • I have made new presets that uses this LoRA and does 8 steps so now you can use them
    • It can go as low as 4 steps but I recommend 8 steps
  • New updated presets are as below

SwarmUI update screenshot 21

  • Audio Only preset is super fast and generates audio directly
    • Audio generation can be used to see if your prompt and duration matching before generating video since it is like real 2x time speed
      • So 1 minute audio generation takes like 30 seconds on RTX 5090 therefore use audio generation preset to see if your prompt and duration matches - for speaking having videos
      • So quickly iterate, see if your audio accurate, then generate full video
  • Moreover, I recommend you to use 8 steps and 0.4 megapixel resolution to quickly generate your videos, verify they are accurate, then move full generation like 1344x768px and 20 steps high quality
    • This works great
  • Add references and write prompt field improved so that now you can add videos and audios with trim as you wish as below (trimming is optional)
    • For @video1's soundtrack, type ; @audio1 is the first standalone audio file
    • According to the above rule, MiniMax_H3_Enchance_Prompt_Feed_For_LLMs.txt file improved and literally working amazing
    • It is located inside Prompt_Generate_LTX_MiniMax_And_Presets_How_To_Use folder inside zip file
  • Trim interface supports both Audio and Video trimming

SwarmUI update screenshot 22

SwarmUI update screenshot 23

  • ComfyUI presets have amazing batch folder processing read changelogs : https://www.patreon.com/SECourses/posts/comfyui-auto-2-105023709
  • For SwarmUI, you can use Wildcards as a batch processing
    • To process them with order not randomly, enable Display Advanced Options
    • Then find Swarm Internal and change Wildcard Seed Behaviour to index
    • Set regular seed 0 so it will start from first prompt in Wildcard and how many generations you make, it will make with order
🔥 Join developers growing publicly
Share your knowledge, build in public, and grow your developer presence with a global community.

More Posts

I’m a Senior Dev and I’ve Forgotten How to Think Without a Prompt

Karol Modelski - Mar 19

Your AI Doesn't Just Write Tests. It Runs Them Too.

Kevin Martinez - May 12

The Sovereign Vault — A Comprehensive Guide to Protocol-Driven AI

Ken W. Algerverified - Jun 4

Tailwind CSS for Beginners: From Zero to Your First Component

muhammadfarhan.dev - Jul 21

MCP Is the USB-C of AI. So Why Are You Plugging Everything In?

Ken W. Algerverified - Jun 10
chevron_left
1.3k Points141 Badges
Türkiye, Mersinpatreon.com/SECourses
34Posts
6Comments
PhD Computer Engineer and Assistant Professor at Computer Engineering Department

100+ Generative A... Show more

Related Jobs

View all jobs →

Commenters (This Week)

5 comments
2 comments
2 comments

Contribute meaningful comments to climb the leaderboard and earn badges!