Text to videoBlack Forest Labs unveils FLUX 3, a unified multimodal foundation model generating 20-second video with native audio, images, and action prediction for robotics.
API-Only
ComfyUI ecosystem updates — open-source model releases, custom nodes, workflows, and tools for image, video, and audio generation.
Text to videoBlack Forest Labs unveils FLUX 3, a unified multimodal foundation model generating 20-second video with native audio, images, and action prediction for robotics.
API-Only
Text to imageMicrosoft Asia releases Mage-Flow -- compact 4B model for text-to-image and editing, native resolution up to 2048, RL/4-step Turbo variants, MIT license.
MIT
Image to imageDamkohler released JLC Flux2 ControlNet v1.0.0, a custom node bringing FLUX.2-dev Fun ControlNet Union into ComfyUI with multi-branch control, reference conditioning, and in/out-paint.
Oasis Suite v1.5 expands Image Oasis into a three-node pack, adding Video Oasis Viewer for in-node preview and LTX2.3 Oasis for all-in-one LTX 2.3 video generation.
Lightricks releases Clean Plate IC-LoRA for LTX-2.3 — a video-to-video LoRA that removes people and objects from footage and reconstructs the empty background, no mask required.
tarn59 released two style transfer LoRAs for Bernini-R — anime style and real life style — enabling video-to-video style conversion directly in ComfyUI.
Apache-2.0
Nvidia Research releases PiD v1.5 with improved decoding quality for FLUX, FLUX.2, and Qwen-Image — better color, fewer artifacts, and enhanced detail.
DreamFast releases an uncensored Qwen3-VL-4B text encoder for ComfyUI, achieving 100% HarmBench ASR across BF16, FP8, INT8, NVFP4 and MXFP8 formats.
Apache-2.0
Lightricks releases an official LTX-2.3 Foley V2A LoRA for realistic, visually synchronized sound effects — footsteps, impacts, typing, and environmental audio.
ComfyUI v0.28.0 adds native SeedVR2 upscaling, NVIDIA PID 1.5 PixelDiT support, convrot int4 optimizations, new Save 3D and Text Overlay nodes, and partner node updates including sync.so sync-3 and Seedream 5 Pro.
Text to videoTongyi Lab (Alibaba) releases Wan-Dancer-14B, an open-source model that generates 720p/30fps dance videos from music and reference images, supporting Chinese Classical, K-Pop, Street, Tap, and Latin dance genres for minute-long outputs.
Fal releases Ideogram V4 Fast (20-step FP4, 5s inference) and Ideogram V4 Instant (8-step BF16, 2s inference), speed-distilled variants with no-CFG support. Available in int8-convrot format for ComfyUI via Hippotes.
Open-Weights
Cseti releases a novel IC-LoRA for LTX-Video 2.3 that acts as a virtual second camera, letting users re-render video scenes from new camera angles using simple text prompts.
Apache-2.0
Cseti releases a proof-of-concept IC-LoRA for LTX-2.3 that acts as a virtual second camera, letting you change the camera angle of existing video footage using a discrete prompt vocabulary.
LiconStudio releases V2 of the popular Multiple Subject Reference LoRA for LTX-2.3, with improved consistency, stability, and scene logic based on community feedback.
Text to videoAlaya Lab released AlayaWorld, a full-stack open source framework for building interactive generative worlds. It features long-horizon video consistency beyond one minute, real-time camera control, and prompt-driven interaction.
rzgar releases Bernini-R-S2V, a speech-to-video fine-tune that adds Wan2.2 S2V audio-driven lip-sync capabilities to ByteDance's Bernini-R model, with dedicated ComfyUI custom nodes and multiple precision variants.
Apache-2.0
Ant Group's Robbyant open-sources LingBot-Video, a 30B-A3B MoE video generation model achieving SOTA on RBench with Apache 2.0 license.
Apache-2.0
A new LoRA from ostris enables style reference transfer on Krea 2 Turbo, trained on thousands of hand-curated style pairs with AI Toolkit.
Lightricks releases the Cinemagraph IC-LoRA for LTX-2.3, an open-weight adapter that keeps backgrounds frozen while moving only a single element — delivering seamless, natural loops from any still image.
LTX-2 Community License