Skip to content
@QwenAudio

QwenAudio

Open-source speech and audio language models from the QwenAudio Team

Popular repositories Loading

  1. CosyVoice CosyVoice Public

    Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.

    Python 23.5k 2.7k

  2. SenseVoice SenseVoice Public

    Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

    C 9.3k 823

  3. qwen-audio-agent qwen-audio-agent Public

    A realtime voice runtime that keeps Agents talking, working, and present. Real-time Voice Runtime for AI Agents

    JavaScript 2.4k 210

  4. Fun-ASR Fun-ASR Public

    Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp deployment paths.

    C 1.5k 151

  5. ThinkSound ThinkSound Public

    [NeurIPS 2025] PyTorch implementation of [ThinkSound], a unified framework for generating audio from any modality, guided by Chain-of-Thought (CoT) reasoning.

    Python 1.4k 82

  6. Fun-Audio-Chat Fun-Audio-Chat Public

    Fun-Audio-Chat is a Large Audio Language Model built for natural, low-latency voice interactions.

    Python 1k 105

Repositories

Showing 10 of 19 repositories
  • QwenAudio/qwen-audio-toolkits's past year of commit activity
    Rust 32 Apache-2.0 3 0 0 Updated Sep 10, 2026
  • Fun-ASR Public

    Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp deployment paths.

    QwenAudio/Fun-ASR's past year of commit activity
    C 1,531 Apache-2.0 151 6 0 Updated Sep 10, 2026
  • qwen-audio-agent Public

    A realtime voice runtime that keeps Agents talking, working, and present. Real-time Voice Runtime for AI Agents

    QwenAudio/qwen-audio-agent's past year of commit activity
    JavaScript 2,390 Apache-2.0 210 12 2 Updated Sep 10, 2026
  • FunResearch Public

    This repository is maintained by the Speech Team at Alibaba’s Tongyi Lab, serving as an open-source platform for our cutting-edge research in speech, audio, NLP technologies. We believe in accelerating scientific progress through transparent collaboration, and invite the global research community to explore, reproduce, and build upon our work.

    QwenAudio/FunResearch's past year of commit activity
    Python 56 Apache-2.0 6 2 0 Updated Sep 8, 2026
  • QwenAudio/QwenAudio.github.io's past year of commit activity
    HTML 1 MIT 1 0 1 Updated Sep 7, 2026
  • SenseVoice Public

    Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

    QwenAudio/SenseVoice's past year of commit activity
    C 9,283 MIT 823 7 0 Updated Sep 5, 2026
  • QwenAudio/FunAudioLLM.github.io's past year of commit activity
    HTML 61 MIT 11 0 0 Updated Jul 25, 2026
  • llama-index-readers-funasr Public

    FunASR (SenseVoice/Paraformer/Fun-ASR-Nano) audio reader for LlamaIndex

    QwenAudio/llama-index-readers-funasr's past year of commit activity
    Python 2 MIT 0 0 0 Updated Jun 17, 2026
  • langchain-funasr Public

    FunASR (SenseVoice/Paraformer/Fun-ASR-Nano) speech-to-text integration for LangChain

    QwenAudio/langchain-funasr's past year of commit activity
    Python 1 0 0 0 Updated Jun 17, 2026
  • funasr-haystack Public archive

    FunASR (SenseVoice/Paraformer) speech-to-text integration for Haystack

    QwenAudio/funasr-haystack's past year of commit activity
    2 0 0 0 Updated Jun 17, 2026