NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 is a new large language model that packs 550 billion total parameters while activating only 55 billion during use. It combines Mamba-2, mixture-of-experts (MoE), and attention layers into a […]
News
Boson AI has released higgs-audio-v3-tts-4b, a 4-billion-parameter text-to-speech model designed specifically for conversational voice AI. Rather than simply reading text aloud, the model produces expressive speech with emotional tone, natural […]
The Nemotron-3.5-ASR-Streaming-0.6b model is NVIDIA’s latest open speech recognition release, designed to transcribe audio in real time across 40 language-locales from a single model. It can handle both low-latency streaming […]
Eclipse-Senpai has released KeyLM-75M, a compact 75 million-parameter language model trained from scratch on roughly 18 billion tokens. This base text-completion model outputs plain text completions and is accompanied by […]
Hcompany has released Holo-3.1-0.8B, the smallest model in a fresh family of vision-language models built to drive computer use agents. The release expands automation capabilities beyond web browsers and desktops […]
Holo-3.1-35B-A3B is the largest model in a new family of vision-language agents that can see, understand, and control computer interfaces across web browsers, desktops, and now mobile devices. It automates […]
Nex-N2-mini is a new open-source AI model designed to handle complex, multi-step tasks by turning its own reasoning into real actions. It is the smaller, more efficient sibling of the […]
Nex-N2-Pro is a new open-source AI model designed to handle complex, real-world agentic tasks. It unifies reasoning, tool use, and code execution into a single continuous workflow called Agentic Thinking. […]
KVarN is a new KV-cache quantization method that expands context capacity for large language models. It compresses keys and values to 4-bit and 2-bit without calibration, yielding up to 3-5x […]