Generative Refinement Networks, or GRN, offers a new method for creating digital images and video. The approach replaces standard diffusion techniques with progressive visual refinement and adaptive computing. Built by […]
News
DisCa is a new acceleration framework designed to speed up AI video generation models. It reduces processing time by 11.8 times without sacrificing visual clarity. Tencent researchers who also made […]
ControlFoley transforms video clips into synchronized soundtracks by combining visual scenes, written descriptions, and existing audio samples into a single generation system. This new framework produces matching sound effects and […]
Combining specialized AI adapters into one system now works best when harmful layers are removed first. The ENMP-LoRAMerging project scans multiple adaptation files, identifies components that hurt overall accuracy, and […]
Yovecent has released UDM-GRPO, an open-source framework that combines uniform discrete diffusion with reinforcement learning for text-to-image generation. The system stabilizes training and improves output quality by treating the fully […]
Tencent researchers recently published MegaStyle, a system designed to automate the creation of visual style libraries. The pipeline translates text descriptions into images that share matching artistic qualities while keeping […]
Tstars-VTON is an open evaluation dataset designed to test virtual try-on models under realistic shopping conditions. It contains 1,780 image pairs covering layered clothing, footwear, and accessories across dozens of […]
SmartPhotoCrafter is an open-source framework that edits photographs without requiring manual prompts. The system automatically spots visual flaws, plans specific improvements, and applies corrections in a single continuous workflow. Researchers […]
AnyRecon turns scattered photographs into complete three-dimensional scenes using a video-based artificial intelligence system. The framework processes inputs in any order without needing precise spacing between camera angles. OpenImagingLab built […]