Best Open Speech & Audio Models
A linked shortlist of open speech & audio models ranked by license clarity, popularity, and neutral trust metadata. Canonical URL: https://huggingbay.xyz/answers/best-open-speech-audio-models.
Best Open Speech & Audio Models is a citation-ready Hugging Bay answer pack backed by live catalog rows. Representative rows: handy-computer/Voxtral-Mini-4B-Realtime-2602-gguf, facebook/seamless-m4t-v2-large, Abiray/LTX-2.5-Distilled-GGUF, Oriserve/Whisper-Hindi2Hinglish-Prime, FunAudioLLM/Fun-ASR-Nano-2512-GGUF. Use each canonical citation URL plus metadata/card endpoints for current license, trust, hosting, and download state. Canonical URL: https://huggingbay.xyz/answers/best-open-speech-audio-models.
Question: What are the best open speech & audio models?
A linked shortlist of open speech & audio models ranked by license clarity, popularity, and neutral trust metadata.
Answer pack JSON Citation pack License guide
- handy-computer/Voxtral-Mini-4B-Realtime-2602-gguf: Model from Hugging Face for automatic speech recognition: handy-computer/Voxtral-Mini-4B-Realtime-2602-gguf · transcribe.cpp
- facebook/seamless-m4t-v2-large: Model from Hugging Face for automatic speech recognition: facebook/seamless-m4t-v2-large · transformers
- Abiray/LTX-2.5-Distilled-GGUF: Model from Hugging Face for text to video: Abiray/LTX-2.5-Distilled-GGUF · unknown
- Oriserve/Whisper-Hindi2Hinglish-Prime: Model from Hugging Face for automatic speech recognition: Oriserve/Whisper-Hindi2Hinglish-Prime · transformers
- FunAudioLLM/Fun-ASR-Nano-2512-GGUF: Model from Hugging Face for automatic speech recognition: FunAudioLLM/Fun-ASR-Nano-2512-GGUF · audio.cpp
- handy-computer/Fun-ASR-MLT-Nano-2512-gguf: Model from Hugging Face for automatic speech recognition: handy-computer/Fun-ASR-MLT-Nano-2512-gguf · transcribe.cpp
- handy-computer/SenseVoiceSmall-gguf: Model from Hugging Face for automatic speech recognition: handy-computer/SenseVoiceSmall-gguf · transcribe.cpp
- handy-computer/granite-4.0-1b-speech-gguf: Model from Hugging Face for automatic speech recognition: handy-computer/granite-4.0-1b-speech-gguf · transcribe.cpp
- Edge0/ARK-ASR-3B: Model from Hugging Face for automatic speech recognition: Edge0/ARK-ASR-3B · transformers
- handy-computer/granite-speech-4.1-2b-plus-gguf: Model from Hugging Face for automatic speech recognition: handy-computer/granite-speech-4.1-2b-plus-gguf · transcribe.cpp
- seanghay/Qwen3-ASR-0.6B-Khmer: Model from Hugging Face for automatic speech recognition: seanghay/Qwen3-ASR-0.6B-Khmer · transformers
- handy-computer/moss-transcribe-diarize-gguf: Model from Hugging Face for automatic speech recognition: handy-computer/moss-transcribe-diarize-gguf · transcribe.cpp
Open interactive answer pack
Updated 2026-09-22 from the live catalog.