GGUF
vokra

audioldm2 (Vokra GGUF)

Converted to the Vokra GGUF format for Vokra, a zero-dependency speech-AI inference runtime.

This is a conversion, not a new model. The weights are the upstream ones; Vokra re-packages them so its runtime can memory-map them directly. Credit for the model belongs upstream β€” see Source below.

Files

File Size SHA-256
audioldm2.gguf 4266.4 MB d02dd97944b05bd27ea17eb0c31034c27992c0f60e642cffef28047a436cb5f1

Usage

# Download (any HTTP client works β€” the file is a plain GGUF)
curl -L -o audioldm2.gguf \
  https://huggingface.co/vokra/audioldm2/resolve/main/audioldm2.gguf
vokra-cli run --model audioldm2.gguf --input input.wav

Provenance

Field Value
Architecture audioldm2
Tensors 2827
Upstream source cvssp/audioldm2 (Liu et al. 2024 arXiv:2308.05734 text-to-audio LDM, cc-by-nc-sa-4.0)
Upstream licence cc-by-nc-sa-4.0
Licence class non-commercial-share-alike
Registry model id audioldm2
Vokra GGUF schema 1
Converted by vokra-core 0.1.0-alpha.0

Every row above is read out of this file's own vokra.* metadata, so the card cannot claim something the artifact does not carry.

Licence

The weights are distributed under cc-by-nc-sa-4.0, unchanged from upstream. Conversion does not alter the licence, and your obligations run to the upstream author.

⚠️ Share-alike / copyleft β€” this licence travels with the file

This weight is cc-by-nc-sa-4.0, and that licence is not discharged by attribution alone. It attaches to derivatives.

  • This GGUF is a conversion of an upstream weight, so it is itself cc-by-nc-sa-4.0 β€” not Apache-2.0, and not covered by Vokra's own licence.
  • Anything you derive from it (a fine-tune, a re-quantisation, a further format conversion) carries the same licence.
  • Vokra's runtime is Apache-2.0. Loading this weight does not change that, because these licences restrict the terms of redistribution, not use. Shipping the weight onward is what carries the obligation.

β›” Non-commercial β€” you may not use this weight commercially

The upstream weight is cc-by-nc-sa-4.0. It is republished here so the model can be evaluated and used for research, and the licence is unchanged by conversion.

  • Do not use this in a commercial product or service. That restriction is upstream's, not Vokra's, and Vokra cannot waive it.
  • Vokra's engine is Apache-2.0 and imposes no such limit β€” the limit is on this weight. Other models in this organisation are permissively licensed; check each one's card.
  • Vokra's runtime refuses to load a non-commercial weight unless an explicit research flag is set, so this restriction is enforced at load time rather than left to the reader.

Verifying this file

shasum -a 256 audioldm2.gguf
# expect: d02dd97944b05bd27ea17eb0c31034c27992c0f60e642cffef28047a436cb5f1
Downloads last month
-
GGUF
Model size
1B params
Architecture
Hardware compatibility
Log In to add your hardware

We're not able to determine the quantization variants.

Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Paper for vokra/audioldm2