Flux_ID_Adjuster_V2 is a new ComfyUI node that improves identity consistency and realism in Flux.2 Klein 9B images. It tackles the common waxy skin problem by letting users control how reference […]
News
Cosmos3-Super-Text2Image is a new open-source image generation model from NVIDIA that creates high-fidelity pictures directly from written prompts. It is a text-to-image variant of the broader Cosmos3 platform, which is […]
Cosmos3-Super-Image2Video is a new open-source AI model that converts a single image into a short video clip guided by a text description. Released by NVIDIA, the 64-billion-parameter tool along with […]
ByteDance has released Bernini, an open-source framework that unifies video generation and editing through a semantic planning approach. Instead of controlling pixels directly, the system uses a multimodal large language […]
The Flux-2-Klein-9B-Schematic-Lora release offers a set of six LoRA adapters that reframe common computer vision tasks as simple image-editing jobs. Each adapter produces a schematic RGB output—like a depth map […]
Nvidia has released Cosmos3-Nano, a 16-billion-parameter omnimodal model that turns text, images, video, audio, or action data into dynamic video with synced sound, reasoning text, or robot movement commands. The […]
The new Bonsai-Image-Ternary-4B-Gemlite-2bit model compresses a 4-billion-parameter text-to-image diffusion transformer into just 1.21 GB. It uses ternary weights—each limited to -1, 0, or +1 with shared scaling—to shrink the model […]
The Comfyui-Anima-IPadapter custom node brings IP-Adapter support to the Anima DiT model inside ComfyUI. It lets users inject reference image features directly into the generation process using decoupled cross-attention. This […]
LongLive-RAG is an open-source framework that turns long video generation into a retrieval problem. An autoregressive generator can look back over its own output and pull in the most relevant […]