LFM2.5-Audio-1.5B GGUF

Highest quality quant - original LiquidAI's GGUFs are FP16, which may degrade quality by clipping values. (Or maybe not, but why risk it?)

In this repo:

  • Base model in BF16, and Q6_K converted from BF16

  • Mmproj in BF16

  • Vocoder in BF16

  • TTS tokenizer in BF16 and F32 (original tensors are F32)

  • Combined vocoder+tokenizer in BF16

Base and mmproj converted using base llama.cpp, tokenizer and vocoder converted using the script from https://github.com/ggml-org/llama.cpp/pull/18641

Text and ASR inference works in baseline llama.cpp. For TTS, use their fork.

Downloads last month
117
GGUF
Model size
1B params
Architecture
lfm2
Hardware compatibility
Log In to add your hardware

6-bit

16-bit

32-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Auguments/LiquidAI-LFM2.5-Audio-1.5B-GGUF-BF16

Quantized
(7)
this model