geninhu
/

xls-asr-vi-40h-1B

Automatic Speech Recognition

hf-asr-leaderboard

robust-speech-event

Inference Endpoints

Model card Files Files and versions Metrics Training metrics Community

geninhu commited on Jan 29, 2022

Commit

4fb6aa5

•

1 Parent(s): 48f71da

Update README.md

Files changed (1) hide show

README.md +36 -16

README.md CHANGED Viewed

@@ -1,35 +1,55 @@
 ---
 license: apache-2.0
 tags:
 - automatic-speech-recognition
-- geninhu/fpt-vi
-- generated_from_trainer
 model-index:
 - name: xls-asr-vi-40h-1B
-  results: []
 ---
-<!-- This model card has been generated automatically according to the information the Trainer had access to. You
-should probably proofread and complete it, then remove this comment. -->
 # xls-asr-vi-40h-1B
-This model is a fine-tuned version of [facebook/wav2vec2-xls-r-1b](https://huggingface.co/facebook/wav2vec2-xls-r-1b) on the GENINHU/FPT-VI - NA dataset.
-It achieves the following results on the evaluation set:
-- Loss: 4.1691
-- Wer: 0.4133
-## Model description
-More information needed
-## Intended uses & limitations
-More information needed
-## Training and evaluation data
-More information needed
 ## Training procedure

 ---
 license: apache-2.0
+language:
+- vi
 tags:
 - automatic-speech-recognition
+- robust-speech-event
+- common-voice
 model-index:
 - name: xls-asr-vi-40h-1B
+  results:
+  - task:
+      name: Speech Recognition
+      type: automatic-speech-recognition
+    dataset:
+      name: Common Voice 7.0
+      type: mozilla-foundation/common_voice_7_0
+      args: vi
+    metrics:
+       - name: Test WER
+         type: wer
+         value: 34.210
+       - name: Test CER
+         type: cer
+         value: 19.938
 ---
 # xls-asr-vi-40h-1B
+This model is a fine-tuned version of [facebook/wav2vec2-xls-r-1b](https://huggingface.co/facebook/wav2vec2-xls-r-1b) on the 40 hours of Vietnamese ASR data, including common_voice 7.0 and private dataset.
+### Benchmark WER result:
+| | [VIVOS](https://huggingface.co/datasets/vivos) | [COMMON VOICE 7.0](https://huggingface.co/datasets/mozilla-foundation/common_voice_7_0) |
+|---|---|---|
+|without LM| 25.93 | 34.21 |
+|with 4-grams LM|  | |
+### Benchmark CER result:
+| | [VIVOS](https://huggingface.co/datasets/vivos) | [COMMON VOICE 7.0](https://huggingface.co/datasets/mozilla-foundation/common_voice_7_0) |
+|---|---|---|
+|without LM| 9.243 | 19.938 |
+|with 4-grams LM|  | |
+## Evaluation
+Please use eval.py file to run evaluation
+```python
+python eval_custom.py --model_id geninhu/xls-asr-vi-40h-1B --dataset mozilla-foundation/common_voice_7_0 --config vi --split test --log_outputs
+```
 ## Training procedure