Speech Swift is a comprehensive AI speech toolkit designed specifically for Apple Silicon devices. It allows users to run powerful speech models locally, including tools for speech recognition, text-to-speech synthesis, […]
News
GRM2-3b is a new 3-billion parameter AI model built for long-term reasoning and complex problem-solving. Despite its small size, it competes with much larger models in benchmarks and handles multi-step […]
ID-LoRA LTX2.3 is a new tool that generates talking-head videos with synchronized audio using a reference voice and image. It creates personalized video content where both the visual appearance and […]
SAMA-14B is a new open-source AI model designed for instruction-guided video editing. It allows users to modify videos using text instructions while keeping the original motion and temporal details intact. […]
A new ComfyUI custom node called FLUX.2 Klein LoRA Loader brings architecture-aware loading to the FLUX.2 Klein 9B model. The tool automatically converts diffusers-format LoRAs to native FLUX format while […]
ComfyUI-advanced-model-manager is a custom node that brings model browsing and downloading directly into ComfyUI. Users can search across hundreds of HuggingFace repositories, download files to the correct folders, and manage […]
ImageTagger is a desktop annotation tool designed for managing image and text pairs, specifically built for machine learning dataset curation workflows. The application provides a streamlined interface for teams and […]
The Michael Hafftka Catalog Raisonné is a new open dataset containing approximately 3,800 artworks by a single artist spanning five decades. The collection covers work from the 1970s through 2025 […]
SANA-Video is a new diffusion model designed to create high-quality videos from text prompts. It can generate content up to 2K resolution with minute-long duration while maintaining strong alignment between […]