QwenAudio
★ 23,769CosyVoice
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
HYSEN LABS DIRECTORY
Verified repositories, classifications and analysis from the QwenAudio organization.
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.
A realtime voice runtime that keeps Agents talking, working, and present. Real-time Voice Runtime for AI Agents
Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp deployment paths.