Qwen3 Reranker 8B Q4_K_M GGUF

This repository contains the exact immutable local-model asset used by NovelAide. This GGUF is a NovelAide-built derived artifact produced from the official Safetensors checkpoint. It is not an official Qwen GGUF release.

Provenance

  • Base model: Qwen/Qwen3-Reranker-8B
  • Asset source: Qwen/Qwen3-Reranker-8B
  • Pinned source revision: 77d193c791ed757ca307ee72715aa132723da912
  • Runtime: llama-cpp
  • NovelAide asset ID: qwen3-reranker-8b
  • NovelAide CDN prefix: qwen3-reranker-8b-gguf/456176c855aba6e6

The base model and conversion input are the official Qwen/Qwen3-Reranker-8B Safetensors at immutable revision 77d193c791ed757ca307ee72715aa132723da912.

NovelAide converts that checkpoint to BF16 GGUF and then quantizes it to Q4_K_M with ggml-org/llama.cpp tag b9842, commit 6f4f53f2b7da54fcdbbecaaa734337c337ad6176. This is the llama.cpp revision embedded by the Desktop runtime's node-llama-cpp@3.19.0.

The resulting GGUF is a NovelAide-built derived asset, not an upstream Qwen GGUF. Its immutable content-addressed object prefix is qwen3-reranker-8b-gguf/456176c855aba6e6; it must not overwrite the previously published community-derived artifact.

Files

File Size SHA-256
Qwen3-Reranker-8B-Q4_K_M.gguf 4460.17 MiB 456176c855aba6e6cc23335fbf285bdb41f77f9e4e5eeb2e72077749482c3b89

Usage

Use the GGUF files with a compatible llama.cpp runtime and the ONNX files with a compatible ONNX / Transformers.js runtime. NovelAide pins the exact files and checksums shown above; do not substitute similarly named quantizations.

License and attribution

The model asset follows the upstream apache-2.0 license. Review the linked base model and asset source model cards for their complete terms, limitations, and attribution requirements. NovelAide is not affiliated with or endorsed by the upstream model authors.

Downloads last month
153
GGUF
Model size
8B params
Architecture
qwen3
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for novelaide/Qwen3-Reranker-8B-Q4_K_M-GGUF

Quantized
(48)
this model