GGUF / GGML model mirror

Byte-for-byte copies of four quantised models: two GGUF language models for llama.cpp and two GGML speech recognition models for whisper.cpp.

Nothing here is original work and nothing has been modified. Each file is a copy of the file named in the table below, mirrored so that a fixed set of weights stays available at a stable address. Upstream repositories get reorganised, re-quantised and occasionally deleted; a mirror pinned to a known checksum does not.

If you are choosing models to use, prefer the upstream repositories linked below โ€” they are maintained, this is not.

Files

File Size Parameters Quantisation
Qwen3-1.7B-Q4_K_M.gguf 1.03 GiB 1.7B Q4_K_M
Llama-3.2-3B-Instruct-Q4_K_M.gguf 1.88 GiB 3.2B Q4_K_M
ggml-base.bin 141 MiB 74M float16
ggml-tiny.bin 74 MiB 39M float16

Provenance and checksums

Copied 2026-08-30. Each file was verified against its SHA-256 after download and again after upload.

File Copied from SHA-256
Qwen3-1.7B-Q4_K_M.gguf unsloth/Qwen3-1.7B-GGUF b139949c5bd74937ad8ed8c8cf3d9ffb1e99c866c823204dc42c0d91fa181897
Llama-3.2-3B-Instruct-Q4_K_M.gguf bartowski/Llama-3.2-3B-Instruct-GGUF 6c1a2b41161032677be168d354123594c0e6e67d2b9227c84f296ad037c728ff
ggml-base.bin ggerganov/whisper.cpp 60ed5bc3dd14eea856493d334349b405782ddcaf0028d4b5df4088345fba2efe
ggml-tiny.bin ggerganov/whisper.cpp be07e048e1e599ad46341c8d2a135645097a538221678b7acdd1b1919c6e1b21

Credit for the quantisations and conversions belongs to their publishers: Qwen3 GGUF by unsloth, Llama 3.2 GGUF by bartowski, and the Whisper GGML conversions by Georgi Gerganov. The base weights are by Alibaba Cloud (Qwen3), Meta (Llama 3.2) and OpenAI (Whisper).

Licences

Three different licences apply, each only to its own files. See NOTICE for the attribution notices that must travel with these copies.

  • Qwen3-1.7B-Q4_K_M.gguf โ€” Apache License 2.0, LICENSE-apache-2.0.txt.

  • Llama-3.2-3B-Instruct-Q4_K_M.gguf โ€” Llama 3.2 Community License, LICENSE-llama-3.2.txt. Use is also governed by Meta's Acceptable Use Policy at https://www.llama.com/llama3_2/use-policy, which the licence incorporates by reference.

    Built with Llama. Llama 3.2 is licensed under the Llama 3.2 Community License, Copyright ยฉ Meta Platforms, Inc. All Rights Reserved.

  • ggml-base.bin, ggml-tiny.bin โ€” MIT, LICENSE-mit.txt.

Downloads last month
-
GGUF
Model size
3B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support