Members
-
CosyVoice
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
Python ★ 23k 2mo agoExplain → -
SenseVoice
Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.
C ★ 9.1k 1d agoExplain → -
qwen-audio-agent
A realtime voice runtime that keeps Agents talking, working, and present. Real-time Voice Runtime for AI Agents
JavaScript ★ 2.1k 4h agoExplain → -
Fun-ASR
Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp runtimes.
C ★ 1.5k 20d agoExplain → -
ThinkSound
[NeurIPS 2025] PyTorch implementation of [ThinkSound], a unified framework for generating audio from any modality, guided by Chain-of-Thought (CoT) reasoning.
Python ★ 1.4k 4mo agoExplain → -
Fun-Audio-Chat
Fun-Audio-Chat is a Large Audio Language Model built for natural, low-latency voice interactions.
Python ★ 990 5mo agoExplain → -
FunCineForge
No description.
Python ★ 446 4mo agoExplain → -
FunAudioLLM-APP
No description.
Python ★ 386 2y agoExplain → -
CV3-Eval
No description.
Python ★ 192 11mo agoExplain → -
FunAudioLLM.github.io
No description.
HTML ★ 61 20d agoExplain → -
MME-Emotion
Official repository for the paper “MME-Emotion: A Holistic Evaluation Benchmark for Emotional Intelligence in Multimodal Large Language Models”
Python ★ 50 6mo agoExplain → -
FunResearch
This repository is maintained by the Speech Team at Alibaba’s Tongyi Lab, serving as an open-source platform for our cutting-edge research in speech, audio, NLP technologies. We believe in accelerating scientific progress through transparent collaboration, and invite the global research community to explore, reproduce, and build upon our work.
Python ★ 47 10d agoExplain → -
qwen-audio-toolkits
No description.
TypeScript ★ 27 1d agoExplain → -
FunMusic
A fundamental toolkit designed for music, song, and audio generation
Python ★ 16 1y agoExplain → -
OmniAudio
No description.
Python ★ 11 1y agoExplain → -
llama-index-readers-funasr
FunASR (SenseVoice/Paraformer/Fun-ASR-Nano) audio reader for LlamaIndex
Python ★ 2 1mo agoExplain → -
funasr-haystack ▣
FunASR (SenseVoice/Paraformer) speech-to-text integration for Haystack
★ 2 1mo agoExplain → -
QwenAudio.github.io
No description.
HTML ★ 0 10d agoExplain → -
langchain-funasr
FunASR (SenseVoice/Paraformer/Fun-ASR-Nano) speech-to-text integration for LangChain
Python ★ 0 1mo agoExplain →
No repos match these filters.