Qwen3 ASR 0.6B β Backpack Voice Package
π Backpack Verified
Compact multilingual ASR with language identification. This package stages immutable upstream artifacts for Backpack's voice runtime layer. It does not replace the chat model selected by the user.
Package
| Field | Value |
|---|---|
| Capability | speech-to-text |
| Input | audio |
| Output | text |
| Runtime | qwen-asr |
| Configured runtime revision | qwen-asr==0.0.6 |
| Format | safetensors | | Precision | BF16 | | Package size | 1.8 GiB | | Recommended RAM | 5.0 GB | | Languages | multilingual |
Validation status
The packager verified the immutable revision, selected-file inventory, non-empty files, hashes, configuration JSON, and the primary artifact container/header. It also loaded the package and passed deterministic audio inference with the configured runtime.
| Integrity | Metadata | Runtime load | Audio inference | Tokenizer |
|---|---|---|---|---|
| passed | passed | passed | passed | passed |
Run with qwen-asr
Install qwen-asr==0.0.6, then load this repository path with
Qwen3ASRModel.from_pretrained(...) and call transcribe(audio=...).
Provenance
Upstream: Qwen/Qwen3-ASR-0.6B
Immutable revision:
5eb144179a02acc5e5ba31e748d22b0cf3e303b0License:
apache-2.0Backpack copied the selected upstream artifacts without modifying model weights.
Backpack did not train this model and does not claim ownership of it.
Files and checksums
chat_template.jsonβ 1.1 KiB β75a8cfca24f00de72d796fbfed6858fc9614ef3dabd8696684cc3bc03a9c58ffconfig.jsonβ 6.0 KiB β76d3ae4601ce939830b2517f4a6cadb86cc51316c3900af6b020b051c21a478cgeneration_config.jsonβ 142.0 B β1da527824d81e07118facff437e03f2e24a23311e3bdeb2368973fe77e5f275cmerges.txtβ 1.6 MiB β8831e4f1a044471340f7c0a83d7bd71306a5b867e95fd870f74d0c5308a904d5model.safetensorsβ 1.7 GiB β79d6cbd4c98c7bbffe9db2edac07f56cd6637d0d5944b27f6c2b8353840323eapreprocessor_config.jsonβ 330.0 B β45e120a4eda2c20c5d7f2ea9354e63536bf35e27aa573fb7cdf78017b378770dtokenizer_config.jsonβ 12.2 KiB β4942d005604266809309cabc9f4e9cb89ce855d59b14681fdc0e1cc62ea26c4cvocab.jsonβ 2.6 MiB βca10d7e9fb3ed18575dd1e277a2579c16d108e32f27439684afa0e10b1440910
Review the upstream model card and license before use or redistribution. Speech systems can mis-transcribe, synthesize misleading content, or behave differently across languages and accents.
- Downloads last month
- -
Model tree for backpack-run/Qwen3-ASR-0.6B-Backpack-ASR
Base model
Qwen/Qwen3-ASR-0.6B