Text Ranking
Transformers
Safetensors
multilingual
qwen3
text-generation
reranker
nvfp4
fp4
compressed-tensors
llm-compressor
vllm
Eval Results (legacy)
8-bit precision
Instructions to use KoliaNik/Qwen3-Reranker-4B-NVFP4A16 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use KoliaNik/Qwen3-Reranker-4B-NVFP4A16 with Transformers:
# Load model directly from transformers import AutoTokenizer, AutoModelForCausalLM tokenizer = AutoTokenizer.from_pretrained("KoliaNik/Qwen3-Reranker-4B-NVFP4A16") model = AutoModelForCausalLM.from_pretrained("KoliaNik/Qwen3-Reranker-4B-NVFP4A16", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Welcome to the community
The community tab is the place to discuss and collaborate with the HF community!