Qwen3 ASR 0.6B β€” Backpack Voice Package

πŸŽ’ Backpack Verified

Compact multilingual ASR with language identification. This package stages immutable upstream artifacts for Backpack's voice runtime layer. It does not replace the chat model selected by the user.

Package

Field Value
Capability speech-to-text
Input audio
Output text
Runtime qwen-asr
Configured runtime revision qwen-asr==0.0.6

| Format | safetensors | | Precision | BF16 | | Package size | 1.8 GiB | | Recommended RAM | 5.0 GB | | Languages | multilingual |

Validation status

The packager verified the immutable revision, selected-file inventory, non-empty files, hashes, configuration JSON, and the primary artifact container/header. It also loaded the package and passed deterministic audio inference with the configured runtime.

Integrity Metadata Runtime load Audio inference Tokenizer
passed passed passed passed passed

Run with qwen-asr

Install qwen-asr==0.0.6, then load this repository path with Qwen3ASRModel.from_pretrained(...) and call transcribe(audio=...).

Provenance

  • Upstream: Qwen/Qwen3-ASR-0.6B

  • Immutable revision: 5eb144179a02acc5e5ba31e748d22b0cf3e303b0

  • License: apache-2.0

  • Backpack copied the selected upstream artifacts without modifying model weights.

  • Backpack did not train this model and does not claim ownership of it.

Files and checksums

  • chat_template.json β€” 1.1 KiB β€” 75a8cfca24f00de72d796fbfed6858fc9614ef3dabd8696684cc3bc03a9c58ff
  • config.json β€” 6.0 KiB β€” 76d3ae4601ce939830b2517f4a6cadb86cc51316c3900af6b020b051c21a478c
  • generation_config.json β€” 142.0 B β€” 1da527824d81e07118facff437e03f2e24a23311e3bdeb2368973fe77e5f275c
  • merges.txt β€” 1.6 MiB β€” 8831e4f1a044471340f7c0a83d7bd71306a5b867e95fd870f74d0c5308a904d5
  • model.safetensors β€” 1.7 GiB β€” 79d6cbd4c98c7bbffe9db2edac07f56cd6637d0d5944b27f6c2b8353840323ea
  • preprocessor_config.json β€” 330.0 B β€” 45e120a4eda2c20c5d7f2ea9354e63536bf35e27aa573fb7cdf78017b378770d
  • tokenizer_config.json β€” 12.2 KiB β€” 4942d005604266809309cabc9f4e9cb89ce855d59b14681fdc0e1cc62ea26c4c
  • vocab.json β€” 2.6 MiB β€” ca10d7e9fb3ed18575dd1e277a2579c16d108e32f27439684afa0e10b1440910

Review the upstream model card and license before use or redistribution. Speech systems can mis-transcribe, synthesize misleading content, or behave differently across languages and accents.

Downloads last month
-
Safetensors
Model size
0.9B params
Tensor type
BF16
Β·
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for backpack-run/Qwen3-ASR-0.6B-Backpack-ASR

Finetuned
(50)
this model