Do not use this.

This is a failed post-training experiment on Clyrai_Lucius-1.3B-Base.

The base was already underfed (~1B tokens on 1.34B dense weights). This fork tried to teach thinking tokens and effort tags (<|effort_low|>, <|effort_medium|>, <|effort_high|>, scratchpad delimiters) via SFT → GRPO → fusion → DPO.

It did not work. Format without knowledge is theatre. The model is not a product, not an assistant, not a reasoning system, and not something to ship.

The model Clyrai stands behind for this line is the base:
https://huggingface.co/saishshinde15/Clyrai_Lucius-1.3B-Base

The sparse successor is Maximus-MoE.

Weights stay up so the experiment is inspectable. That is the only reason.


The work is ours.

Research and personal evaluation, with credit. Nothing else.

Clyrai Sovereign Research License 1.0

Title stays with Clyrai.

Downloads last month
616
Safetensors
Model size
1B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for saishshinde15/Clyrai_Lucius-1.3B

Finetuned
(1)
this model
Quantizations
1 model