Skip to content
@QwenAudio

QwenAudio

Open-source speech and audio language models from the QwenAudio Team

Popular repositories Loading

  1. CosyVoice CosyVoice Public

    Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.

    Python 23.8k 2.7k

  2. SenseVoice SenseVoice Public

    Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

    C 9.4k 829

  3. qwen-audio-agent qwen-audio-agent Public

    A realtime voice runtime that keeps Agents talking, working, and present. Real-time Voice Runtime for AI Agents

    JavaScript 2.8k 265

  4. Fun-ASR Fun-ASR Public

    Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp deployment paths.

    C 1.6k 152

  5. ThinkSound ThinkSound Public

    [NeurIPS 2025] PyTorch implementation of [ThinkSound], a unified framework for generating audio from any modality, guided by Chain-of-Thought (CoT) reasoning.

    Python 1.4k 82

  6. Fun-Audio-Chat Fun-Audio-Chat Public

    Fun-Audio-Chat is a Large Audio Language Model built for natural, low-latency voice interactions.

    Python 1k 105

Repositories

Showing 10 of 20 repositories
  • FunResearch Public

    This repository is maintained by the Qwen Audio Team at Alibaba Group, serving as an open-source platform for our cutting-edge research in speech, audio, NLP technologies. We believe in accelerating scientific progress through transparent collaboration, and invite the global research community to explore, reproduce, and build upon our work.

    QwenAudio/FunResearch's past year of commit activity
    Python 61 Apache-2.0 6 2 0 Updated Sep 24, 2026
  • qwen-audio-agent Public

    A realtime voice runtime that keeps Agents talking, working, and present. Real-time Voice Runtime for AI Agents

    QwenAudio/qwen-audio-agent's past year of commit activity
    JavaScript 2,750 Apache-2.0 265 3 7 Updated Sep 23, 2026
  • qwen-audio-toolkits Public

    Local-first desktop app for audio AI: conversational agent + on-demand open-source model store.

    QwenAudio/qwen-audio-toolkits's past year of commit activity
    Rust 36 Apache-2.0 3 0 0 Updated Sep 23, 2026
  • QwenAudio/QwenAudio.github.io's past year of commit activity
    HTML 1 MIT 2 0 1 Updated Sep 23, 2026
  • SenseVoice Public

    Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

    QwenAudio/SenseVoice's past year of commit activity
    C 9,373 MIT 829 7 3 Updated Sep 22, 2026
  • Fun-Audio-Chat Public

    Fun-Audio-Chat is a Large Audio Language Model built for natural, low-latency voice interactions.

    QwenAudio/Fun-Audio-Chat's past year of commit activity
    Python 1,008 Apache-2.0 105 16 3 Updated Sep 21, 2026
  • Fun-ASR Public

    Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp deployment paths.

    QwenAudio/Fun-ASR's past year of commit activity
    C 1,552 Apache-2.0 152 7 0 Updated Sep 10, 2026
  • alibabacloud-bailian-speech-demo Public Forked from aliyun/alibabacloud-bailian-speech-demo

    Sample Code Repository for the AlibabaCloud Bailian Speech SDK

    QwenAudio/alibabacloud-bailian-speech-demo's past year of commit activity
    1 MIT 56 0 0 Updated Sep 9, 2026
  • QwenAudio/FunAudioLLM.github.io's past year of commit activity
    HTML 61 MIT 11 0 0 Updated Jul 24, 2026
  • llama-index-readers-funasr Public

    FunASR (SenseVoice/Paraformer/Fun-ASR-Nano) audio reader for LlamaIndex

    QwenAudio/llama-index-readers-funasr's past year of commit activity
    Python 2 MIT 0 0 0 Updated Jun 17, 2026

Top languages

Loading…