Trending Model:#1Inklingthinkingmachines⬇13kTrending Model:#2Ternary-Bonsai-27B-ggufprism-ml⬇339kTrending Model:#3Bonsai-27B-ggufprism-ml⬇1263kTrending Model:#4Unlimited-OCRbaidu⬇2123kTrending Model:#5GLM-5.2zai-org⬇532kTrending Model:#6Qwythos-9B-Claude-Mythos-5-1M-GGUFempero-ai⬇2117kTrending Model:#7krea2-identity-editconradlocke⬇0kTrending Model:#8Qwen3.6-35B-A3B-Uncensored-HauhauCS-AggressiveHauhauCS⬇2007kTrending Model:#9Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUFDavidAU⬇17kTrending Model:#10OvisOCR2ATH-MaaS⬇15kTrending Model:#1Inklingthinkingmachines⬇13kTrending Model:#2Ternary-Bonsai-27B-ggufprism-ml⬇339kTrending Model:#3Bonsai-27B-ggufprism-ml⬇1263kTrending Model:#4Unlimited-OCRbaidu⬇2123kTrending Model:#5GLM-5.2zai-org⬇532kTrending Model:#6Qwythos-9B-Claude-Mythos-5-1M-GGUFempero-ai⬇2117kTrending Model:#7krea2-identity-editconradlocke⬇0kTrending Model:#8Qwen3.6-35B-A3B-Uncensored-HauhauCS-AggressiveHauhauCS⬇2007kTrending Model:#9Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUFDavidAU⬇17kTrending Model:#10OvisOCR2ATH-MaaS⬇15k

News

June 24, 2026
Qwopus3.6-27B-Coder-MTP-GGUF Speeds Local Code Agents By 1.66x

Qwopus3.6-27B-Coder-MTP-GGUF is a new quantized coding model designed for fast, local agentic software development. It represents a specialized version of the Qwopus3.6-27B-Coder, packaged as a GGUF file for efficient single-GPU […]

Read More
June 24, 2026
OBLITERATUS Sculpts Gemma-4-12B-OBLITERATED A Fully Uncensored Brain With Zero Smarts Lost

Gemma-4-12B-OBLITERATED is the first language model to completely remove built-in safety refusals with absolutely zero loss in benchmark performance. This modified version of Google’s Gemma 4 12B scores identically to […]

Read More
June 24, 2026
North-Mini-Code-1.0 Materializes As An Open Agentic Coding Engine

North-Mini-Code-1.0 is a new open-source AI model from CohereLabs built specifically for code generation, autonomous software engineering, and terminal tasks. It packs a 30-billion-parameter architecture but only activates 3 billion […]

Read More
June 24, 2026
MiniMax-M3 Handles 1M Tokens Across Text Images And Video Natively

MiniMax-M3 is a new native multimodal AI model from MiniMaxAI that processes text, images, and video with a 1 million token context window. The model contains about 428 billion total […]

Read More
June 24, 2026
Moonshot AI Drops Kimi-K2.7-Code with 30% Thinking Token Cut

Moonshot AI has released Kimi-K2.7-Code, an open-source coding agentic model that significantly upgrades long-horizon software engineering performance. It builds directly on the Kimi K2.6 architecture while cutting thinking-token usage by […]

Read More
June 24, 2026
Google Forges Diffusiongemma-26B-A4B-It A Diffusion Model That Denoises Entire Text Blocks At Once

Google has released a new open-weights AI model called Diffusiongemma-26B-A4B-it, which uses a unique method to generate text significantly faster than traditional models. Instead of creating text one token at […]

Read More
June 24, 2026
Anvil Forges Transparent AI Fleets On Your Hardware With Plain Files

Anvil is a new open-source model runner that wraps llama.cpp, providing transparent controls and fleet management for local AI on private hardware. It stores models as plain GGUF files you […]

Read More
June 22, 2026
MindLab Research Serves Up Macaron-V1-Preview-749B Personal AI Agent

MindLab Research has released Macaron-V1-Preview-749B, a 749-billion-parameter AI model built to serve as a personal agent that can use tools and generate dynamic user interfaces. The system combines a massive […]

Read More
June 22, 2026
Give Your AI Server A Checkup With Vllm-Doctor

Vllm-doctor is a new command-line tool that diagnoses performance bottlenecks in vLLM inference servers by reading live metrics. It turns raw data into clear explanations of what is wrong, why […]

Read More
1 10 11 12 13 14 81