Qwen-AgentWorld is a new language world model designed to simulate digital environments for AI agents. It predicts what happens next when an AI takes an action in text or visual […]
News
Qwythos-9B-Claude-Mythos-5-1M is a full-parameter reasoning model built on a deeply uncensored base. It processes up to one million tokens of text at once for massive code analysis and research. The […]
DeepSeek-v4-Fable is an autonomous agent engineered for offensive security research. This release operates as a specialized tool for solving challenges and planning exploits in controlled environments. It acts as a […]
Baidu has introduced Unlimited-OCR, a new tool designed to read and transcribe long documents without losing speed. This model processes dozens of pages in a single pass by maintaining a […]
ComfyUI-Yedp-UV-Painter is an experimental, non-destructive 3D-to-2D AI texturing node for ComfyUI. It bridges professional 3D pipelines with AI image generation directly inside the browser interface. Users can surgically texture complex […]
Supra-A2A-Nano-Exp is an experimental proof-of-concept any-to-any model that processes text, images, and video using a single system. It translates visual inputs into a small set of learned codes and treats […]
MiMo-Audio-7B-Instruct is a new audio language model designed to understand and generate sound based on simple instructions. It learns from a massive amount of audio data to perform tasks like […]
The new release called MiMo-Audio-7B-Base is an open-source audio language model designed to learn new tasks from just a few examples. It processes over one hundred million hours of audio […]
Boogu-Image-0.1-Edit is an open-source image editing and transformation model that works alongside a broader multimodal generation family. This release allows users to perform complex editing tasks like object insertion, background […]