qwen3-4b-sorrel-selfloop-g3-midtrain

Private research artifact — do not redistribute. Anthropic Fellows project on flourishing-framed character training (Wang & Jermyn pitch, 2026-04-22).

  • base model: joshycodes/qwen3-4b-sorrel-selfloop-g2-midtrain @ daffe2475da0
  • steps: midtrain
  • run: qwen3-4b-sorrel-selfloop-g2-midtrain-sorrel-selfloop-b-g2-m-0916-0307 on 2x NVIDIA H200 (RunPod, fellows worker)
  • launcher commit: a0afb77669ae of the flourishing-training repo
  • seed: 20260821
step data revision tokens seen loss
midtrain joshycodes/sorrel-selfloop-corpus (config sorrel-selfloop-b-g2) a6ac5e69dbc6 7,847,936 tokens loss 1.2108 → 1.1221

Hyperparameters:

{
  "midtrain": {
    "lr": 1e-05,
    "seq_len": 4096,
    "micro_batch": 4,
    "grad_accum": 4,
    "epochs": 1.0
  }
}

Full run config in train_run_config.json; evaluate with uv run eval.py --model joshycodes/qwen3-4b-sorrel-selfloop-g3-midtrain --eval all.

Downloads last month
50
Safetensors
Model size
4B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for joshycodes/qwen3-4b-sorrel-selfloop-g3-midtrain

Finetuned
(2)
this model
Finetunes
2 models

Dataset used to train joshycodes/qwen3-4b-sorrel-selfloop-g3-midtrain