Qwentaur-4B-LoRA-r32

Qwen

socius Paper Parameters LoRA Dataset

Qwentaur-4B-LoRA-r32

LoRA adapter for Qwentaur-4B, fine-tuned on the full Psych-101 as part of the LoRA-rank sweep and dataset-size ablation for Small Foundation Models of Human Cognition and Behaviour.

field value
base model unsloth/Qwen3-4B-Base
LoRA rank 32 (alpha = rank, rsLoRA)
data fraction 100% of Psych-101
training 1 epoch, completion-only loss, seed 3407

Load with PEFT on top of unsloth/Qwen3-4B-Base, or evaluate with the project's eval_model.py --backend unsloth.

Downloads last month
6
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for socius/Qwentaur-4B-LoRA-r32

Adapter
(39)
this model

Dataset used to train socius/Qwentaur-4B-LoRA-r32

Collection including socius/Qwentaur-4B-LoRA-r32

Paper for socius/Qwentaur-4B-LoRA-r32