Qwen3-TTS Easy Finetuning is an open-source tool that simplifies the process of training custom voice models. It provides a browser-based interface to manage the entire workflow, from processing raw audio […]
News
ComfyUI-Wan-VACE-Video-Joiner is a workflow tool that automatically stitches multiple video clips together while creating smooth transitions between them. The system uses VACE (Video-to-Video) technology to generate new frames at each […]
ComfyUI-Spectrum-WAN-Proper is a custom node for ComfyUI that speeds up WAN video generation. It uses a technique called Spectrum, which forecasts denoiser features instead of running the full network at […]
Matrix-Game-3.0 is an open-source interactive world model that generates real-time video at 720p resolution and 40 frames per second. It uses a memory-augmented architecture to maintain consistency over long video […]
Qwen-3.5-Abliterated-Comfyui-nvfp4 is a collection of quantized language models designed to function as AI assistants directly within ComfyUI. Developer Winnougan created these models to enable multimodal tasks like image analysis and […]
Z-Image-SAM-ControlNet is a new control model designed to transform segmented images into photorealistic pictures. It functions as a ControlNet for the Tongyi-MAI/Z-Image base model, allowing users to guide image generation […]
LongCat-AudioDiT is a new text-to-speech model that generates high-fidelity audio directly from text inputs. It operates directly on the waveform latent space rather than relying on intermediate acoustic representations like […]
ComfyUI-FBnodes is a collection of custom nodes for ComfyUI that streamlines video workflows and adds utility functions for AI content generation. The extension provides tools for video encoding with codec […]
World Model Bench is a new benchmark that tests whether AI world models can actually think about a scene rather than just generate smooth video. It measures cognitive intelligence through […]