Higgs_v3-TTS-ComfyUI is a new release that brings multilingual conversational text to speech into ComfyUI. It allows users to generate spoken audio in over 100 languages using a local AI model. […]
News
Mordant-12B-Think is a new AI model designed to generate detailed and narratively rich compositions for image generation. It uses chain-of-thought reasoning to analyze spatial relationships and construct deliberate meaning through […]
AudioX-Turbo is a new open source framework that generates audio and music from text, video, and existing audio signals. This release processes your inputs in just four steps to create […]
The newly released Dasheng-Audiogen is an open source artificial intelligence model that creates full audio scenes from text descriptions. Instead of producing just one type of sound, it can blend […]
Difforum is a set of custom nodes for ComfyUI that creates keyframe animations using math expressions, camera moves, and audio reactivity. It updates the classic Deforum workflow to work with […]
The new release of Huggingface-model-filter provides a powerful userscript to sort through Hugging Face models using positive and negative keywords. This tool displays a floating panel on your screen that […]
Texture2albedo-v2 is a new AI tool that extracts pure base color from textured images. It works by removing shadows, reflections, highlights, and specularity from original photos. This process leaves behind […]
The new release called one-node-flux-2-klein wraps the complete FLUX.2 [klein] workflow into a single user interface widget for ComfyUI. Users can access text to image generation, image editing, and painting […]
ComfyUI-Krea2T-Enhancer is a custom node designed to improve how closely Krea2 diffusion models follow text prompts. It works by adjusting the internal conditioning path during the sampling process without altering […]