Skip to content
@QwenAudio

QwenAudio

Open-source speech and audio language models from the QwenAudio Team

Popular repositories Loading

  1. CosyVoice CosyVoice Public

    Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.

    Python 23.8k 2.7k

  2. SenseVoice SenseVoice Public

    Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

    C 9.4k 833

  3. qwen-audio-agent qwen-audio-agent Public

    A realtime voice runtime that keeps Agents talking, working, and present. Real-time Voice Runtime for AI Agents

    JavaScript 2.8k 280

  4. Fun-ASR Fun-ASR Public

    Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp deployment paths.

    C 1.6k 153

  5. ThinkSound ThinkSound Public

    [NeurIPS 2025] PyTorch implementation of [ThinkSound], a unified framework for generating audio from any modality, guided by Chain-of-Thought (CoT) reasoning.

    Python 1.4k 81

  6. Fun-Audio-Chat Fun-Audio-Chat Public

    Fun-Audio-Chat is a Large Audio Language Model built for natural, low-latency voice interactions.

    Python 1k 106

Repositories

Showing 10 of 20 repositories