Trending Model:#1Inklingthinkingmachines⬇16kTrending Model:#2Unlimited-OCRbaidu⬇2237kTrending Model:#3Ternary-Bonsai-27B-ggufprism-ml⬇432kTrending Model:#4Bonsai-27B-ggufprism-ml⬇1405kTrending Model:#5GLM-5.2zai-org⬇545kTrending Model:#6Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUFDavidAU⬇63kTrending Model:#7Laguna-S-2.1poolside⬇3kTrending Model:#8Qwen3.6-35B-A3B-Uncensored-HauhauCS-AggressiveHauhauCS⬇1998kTrending Model:#9krea2-identity-editconradlocke⬇0kTrending Model:#10Qwythos-9B-Claude-Mythos-5-1M-GGUFempero-ai⬇2133kTrending Model:#1Inklingthinkingmachines⬇16kTrending Model:#2Unlimited-OCRbaidu⬇2237kTrending Model:#3Ternary-Bonsai-27B-ggufprism-ml⬇432kTrending Model:#4Bonsai-27B-ggufprism-ml⬇1405kTrending Model:#5GLM-5.2zai-org⬇545kTrending Model:#6Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUFDavidAU⬇63kTrending Model:#7Laguna-S-2.1poolside⬇3kTrending Model:#8Qwen3.6-35B-A3B-Uncensored-HauhauCS-AggressiveHauhauCS⬇1998kTrending Model:#9krea2-identity-editconradlocke⬇0kTrending Model:#10Qwythos-9B-Claude-Mythos-5-1M-GGUFempero-ai⬇2133k

Audio

About audio model releases

Explore the latest open‑source audio and speech AI releases for local use. This archive covers new models and tools for voice cloning, text‑to‑speech, transcription, and music generation.

Latest audio models

April 20, 2026
OpenMOSS-Team Debut MOSS-TTS-Nano-100M Offline Audio Engine

MOSS-TTS-Nano-100M is a lightweight, open-source text-to-speech engine that generates natural audio directly on standard computers. The system converts typed prompts into clear speech while maintaining strict efficiency for daily use. […]

Read More
April 20, 2026
k2-fsa OmniVoice Turns Text To Speech In 600 Languages Offline

OmniVoice is an open source text-to-speech system that converts written words into spoken audio across more than six hundred languages. The software enables instant voice matching and allows users to […]

Read More
April 20, 2026
VoxCPM2 Brings Studio Sound To Local Devices

VoxCPM2 is an open text-to-speech system that generates studio-quality audio from written text. The tool reads words at standard clarity and outputs polished speech at a higher frequency without needing […]

Read More
April 15, 2026
ACE-Step 1.5 XL turns plain text into full songs in eight quick steps

ACE-Step recently published ACE-Step 1.5 XL, an open audio generation model that produces complete music tracks in just eight steps. This streamlined process significantly reduces rendering wait times while preserving […]

Read More
April 7, 2026
Foundation-1 Crafts Structured Loops for Producers

Foundation-1 is a text-to-sample model built for structured music production. It generates tempo-synced, key-aware loops that slot directly into production workflows instead of producing generic audio textures. RoyalCities developed this […]

Read More
April 7, 2026
LongCat-AudioDiT Masters Zero-Shot Voice Cloning with Ease

LongCat-AudioDiT is a new text-to-speech model that generates high-fidelity audio directly from text inputs. It operates directly on the waveform latent space rather than relying on intermediate acoustic representations like […]

Read More
March 30, 2026
PrismAudio Transforms Video into Realistic Soundtracks

PrismAudio is a new framework that generates audio from video using reinforcement learning with Chain-of-Thought (CoT) planning. Developed by the FunAudioLLM team, it breaks down the complex task of video-to-audio […]

Read More
March 26, 2026
Yuriyvnv Refines Dutch Speech Data With WAVe Update

WAVe-1B-Multimodal-NL is a 1 billion parameter model that checks the quality of synthetic speech at the word level. It examines how well spoken audio matches its written transcript, catching errors […]

Read More
March 22, 2026
OpenMOSS MOSS-TTS Speech Studio for home GPUs

MOSS-TTS Family is an open-source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high-fidelity audio generation across complex real-world scenarios, including long-form […]

Read More