whisper-large-v3 (Vokra GGUF)

Speech recognition model, converted to the Vokra GGUF format for Vokra, a zero-dependency speech-AI inference runtime.

This is a conversion, not a new model. The weights are the upstream ones; Vokra re-packages them so its runtime can memory-map them directly. Credit for the model belongs upstream โ€” see Source below.

Files

File Size SHA-256
model.gguf 2944.5 MB 0672cf5550da1eab3209f7eb4c84182be38e1269f88d83eead9e643d758b7fe3

Usage

# Download (any HTTP client works โ€” the file is a plain GGUF)
curl -L -o model.gguf \
  https://huggingface.co/vokra/whisper-large-v3/resolve/main/model.gguf
vokra-cli run --model model.gguf --input speech.wav

# Beam search instead of greedy
vokra-cli run --model model.gguf --input speech.wav --beam-size 5

Prints the transcript. Input must be mono 16 kHz WAV.

Provenance

Field Value
Architecture whisper
Tensors 1259
Upstream source openai/whisper (MIT) โ€” OpenAI official release
Upstream licence MIT
Licence class permissive
Registry model id whisper
Vokra GGUF schema 1
Converted by vokra-core 0.1.0-alpha.0

Every row above is read out of this file's own vokra.* metadata, so the card cannot claim something the artifact does not carry.

Licence

The weights are distributed under MIT, unchanged from upstream. Conversion does not alter the licence, and your obligations run to the upstream author.

Verifying this file

shasum -a 256 model.gguf
# expect: 0672cf5550da1eab3209f7eb4c84182be38e1269f88d83eead9e643d758b7fe3
Downloads last month
154
GGUF
Model size
2B params
Architecture
Hardware compatibility
Log In to add your hardware

We're not able to determine the quantization variants.

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support