Qwen3-ASR Large Voice Model
Qwen3-ASR brings LLM-scale understanding to speech recognition. It excels in noisy audio environments, complex vocabulary, and domain-specific terminologies.
Configuration Example
yaml
ASR-Providers:
qwen:
enable: true
engine: qwen
load:
model_name: Qwen/Qwen3-ASR-1.7B
model_path: models/Qwen3-ASR-1.7B
device: cuda
dtype: float16
max_new_tokens: 256
max_inference_batch_size: 32