# QwenLM/Qwen2-Audio

The official repo of Qwen2-Audio chat & pretrained large audio language model proposed by Alibaba Cloud.

Repository: https://github.com/QwenLM/Qwen2-Audio
Canonical: https://ross.abutalabs.com/products/qwen2-audio
Language: Python
License Family: other
Last push: 2025-04-21T08:50:49+00:00

## Health v2 (maintenance only)
Score: 31/100 (v2, computed 2026-09-02T17:46:02.011165+00:00)
- activity 17, release rhythm 35, longevity 57
- inputs: {"age_days": 800, "days_push": 499, "days_rel": null, "gap_med": null, "n_releases_24m": 0}
- flags: no_releases, no_license
- formula: round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10)

## Adoption (not part of the score)
Stars 2099, forks 169 (observed 2026-08-28T04:06:13.478912+00:00)

## What it is
Official repository for Qwen2-Audio, a 7B-parameter large audio-language model from Alibaba Cloud that accepts audio inputs and responds to speech or text instructions. It provides model weights, inference code, and evaluation for voice chat and audio analysis tasks.

## Use cases
- transcribe speech to text with a large audio model
- build a voice chat assistant that takes audio input
- analyze audio clips with text instructions
- recognize speech emotion from audio
- translate speech to text across languages
- classify vocal sounds in recordings

## When to choose
- you need a single model handling ASR, speech translation, and audio understanding
- you want voice-only interaction without text input
- you need an open 7B audio-language model to fine-tune or self-host

## When to avoid
- you need lightweight or on-device speech recognition
- you only need a dedicated ASR engine like Whisper
- you require a permissive license - the repo lists no license

## Facets
- artifact type: library
- maturity: active
- function: speech-recognition, audio-processing, llm-inference, machine-learning
- domain: speech-processing, large-language-models, artificial-intelligence
- platform: python
- tags: audio-language-model, multimodal, voice-chat, qwen, transformers, model-weights, audio, gpu, linux

## Member repositories
- QwenLM/Qwen2-Audio (main) score 31

## Provenance
- Observed fields: from GitHub, fetched 2026-08-28T04:06:13.478912+00:00.
- Health v2: computed from the inputs above; adoption is never an input.
- Inferred fields (summary, facets, guidance): AI-extracted, prompt v1, taxonomy v1, on 2026-08-30T02:54:22.109987+00:00, confidence not recorded.
  - readme: https://github.com/QwenLM/Qwen2-Audio (fetched 2026-08-28T04:06:13.478912+00:00, sha 0864bb036062)
- Data as of 2026-08-30T08:39:29.467469+00:00.
