Qwen3-ASR Persian Elderly

مدل Qwen3-ASR-1.7B تنظیم‌دقیق‌شده برای بازشناسی گفتار فارسی، با تمرکز بر گفتار سالمندان و مقاومت در برابر تغییرات آکوستیکی و نویز. این مخزن فقط مصنوعات لازم برای inference را منتشر می‌کند و شامل optimizer یا وضعیت ادامه آموزش نیست.

Fine-tuned Qwen3-ASR-1.7B for Persian speech recognition, with emphasis on elderly speech and acoustic robustness. This repository contains inference artifacts only.

Base model

Qwen/Qwen3-ASR-1.7B

Training data

Training used a mixture of Persian public/crawled speech and elderly-speech resources, including Common Voice Persian, Ganjoor-derived data, Filimo-derived data, locally collected elderly Persian speech, and cross-lingual elderly speech data. Audio normalization and probabilistic augmentation such as noise injection, speed perturbation, signal degradation, and pause insertion were used.

Evaluation

The reported aggregate result for the fine-tuned Qwen system is WER 0.31 and CER 0.27. Results depend on the exact normalization and test composition; model comparisons and evaluation code will be published separately.

Intended use and limitations

Intended for Persian ASR research and prototyping. Performance may vary across accents, recording devices, noise conditions, ages, and domains. Outputs can contain omissions or substitutions and should not be treated as authoritative transcripts in safety-critical settings.

Authors

Ali Alvandi — undergraduate project supervised by Hossein Sameti.

Downloads last month
-
Safetensors
Model size
2B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for AliAvd/qwen3-asr-persian-elderly

Finetuned
(100)
this model