Cypher CODE-PRM-8B โ LoRA GGUF (runtime, sin merge)
Adaptador LoRA del especialista destilado CODE-PRM-8B del ecosistema Cypher, convertido a formato GGUF para servir EN RUNTIME con llama.cpp:
llama-server --model Qwen3-8B-Q4_K_M.gguf --lora CODE-PRM-8B-lora.gguf
No requiere merge con la base: se aplica al vuelo. Origen: adapter F32 del repo cypher-CODE-PRM-8B-v8-GGUF (r=16, alpha=32, 7 target_modules).
- Downloads last month
- -
Hardware compatibility
Log In to add your hardware
We're not able to determine the quantization variants.
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐ Ask for provider support