← all repositories

QwenLM/Qwen2-Audio

A large audio-language model from Alibaba's Qwen team capable of accepting speech and audio inputs for conversational interaction and audio analysis.

Qwen2-Audio
Not currently ranked — collecting fresh signals.
star history

Qwen2-Audio is a large-scale audio-language model that processes various audio signals and responds to speech instructions. It supports two interaction modes: free-form voice chat without text input, and audio analysis where users provide audio paired with text queries. The model is released in 7B parameter versions for both pretrained and instruction-tuned variants.

Frequently asked

What is QwenLM/Qwen2-Audio?
A large audio-language model from Alibaba's Qwen team capable of accepting speech and audio inputs for conversational interaction and audio analysis.
Is Qwen2-Audio open source?
Yes — QwenLM/Qwen2-Audio is an open-source project tracked on heatdrop.
What language is Qwen2-Audio written in?
QwenLM/Qwen2-Audio is primarily written in Python.
How popular is Qwen2-Audio?
QwenLM/Qwen2-Audio has 2.1k stars on GitHub.
Where can I find Qwen2-Audio?
QwenLM/Qwen2-Audio is on GitHub at https://github.com/QwenLM/Qwen2-Audio.

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.