Cypher CODE-PRM-8B โ€” LoRA GGUF (runtime, sin merge)

Adaptador LoRA del especialista destilado CODE-PRM-8B del ecosistema Cypher, convertido a formato GGUF para servir EN RUNTIME con llama.cpp:

llama-server --model Qwen3-8B-Q4_K_M.gguf --lora CODE-PRM-8B-lora.gguf

No requiere merge con la base: se aplica al vuelo. Origen: adapter F32 del repo cypher-CODE-PRM-8B-v8-GGUF (r=16, alpha=32, 7 target_modules).

Downloads last month
-
GGUF
Model size
43.6M params
Architecture
qwen3
Hardware compatibility
Log In to add your hardware

We're not able to determine the quantization variants.

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Dennis1315/cypher-code-prm-8b-lora-gguf

Finetuned
Qwen/Qwen3-8B
Adapter
(2166)
this model