The new Bonsai-Image-Ternary-4B-Gemlite-2bit model compresses a 4-billion-parameter text-to-image diffusion transformer into just 1.21 GB. It uses ternary weights—each limited to -1, 0, or +1 with shared scaling—to shrink the model […]
Image
About local image releases
Latest image models
Prism ML has released Bonsai-Image-Binary-4B-Gemlite-1bit, a text-to-image model that packs a full diffusion transformer into just 0.93 GB by using binary weights. It takes the FLUX.2 Klein 4B architecture and […]
Microsoft has released Lens-Turbo, a distilled version of its Lens text-to-image model that can generate high-quality pictures in just four processing steps. Lens is a 3.8-billion-parameter foundational model designed from […]
Microsoft has released Lens, a 3.8-billion-parameter text-to-image model that generates high-quality images with much lower training compute requirements than larger alternatives. It outperforms or matches 6B+ parameter models on standard […]
Nvidia has released PiD, a pixel diffusion decoder that speeds up high‑resolution image generation from latent models. It reformulates the standard decoder as a conditional diffusion model, denoising directly in […]
AsymFLUX.2-klein-9B is an adapter that lets the FLUX.2 klein Base 9B model create images in raw pixel space, bypassing the usual VAE (decoder) step. It uses an asymmetric flow method […]
Pixal3D is a new open-source model that turns a single image into a detailed 3D asset with high fidelity. It goes beyond typical generation methods by creating a direct pixel-to-3D […]
Anima Base v1.0 is a 2 billion parameter text-to-image model built to generate anime-style and other non-photorealistic artwork. It creates illustrations with a focus on anime concepts, characters, and styles, […]
HiDream-O1-Image is an open-source image generation model that creates, edits, and personalizes visuals without relying on separate compression tools. It uses a Pixel-level Unified Transformer to process raw pixels, text, […]