millie-v4-krea2-lora

Krea 2 character LoRA โ€” trigger woman. Trained on Krea 2 RAW, intended for inference on Krea 2 Turbo.

Status: โœ… training complete

Training configuration

Field Value
Base model (train) krea/Krea-2-Raw (raw.safetensors, 12B DiT, undistilled)
Inference target krea/Krea-2-Turbo (8-step distilled) โ€” RAW-train / Turbo-infer
VAE Qwen-Image VAE (qwen_image_vae.safetensors)
Text encoder Qwen3-VL-4B (qwen3vl_4b_bf16.safetensors)
Trainer Musubi Tuner (kohya-ss), networks.lora_krea2
LoRA rank/dim 24, alpha 24, all Linear layers of the DiT
Optimizer adamw8bit, lr 1e-4
Precision bf16, gradient checkpointing, SDPA
Timestep sampling krea2_shift (resolution-aware), weighting_scheme none
Trigger woman (no dedicated trigger token โ€” class word only; identity learned from the dataset (young woman))
Dataset 66 seedream-v5-pro 2k + a few gpt-image Western-animation seductive portrait+medium images + paired natural-language captions; clothing/pose/expression/framing/background described; young woman, no dedicated trigger (class word 'woman'). batch 1, 1024 multi-res buckets, num_repeats 2
Schedule 12 epochs = 1584 steps, save every 2 epochs, seed 42
Hardware 1x NVIDIA H100 80GB (RunPod, EU-NL-1, Musubi volume)

Loss per checkpoint

avr_loss is a rolling average โ€” flow-matching loss is near-flat, so use it as a sanity check and pick the best checkpoint visually.

checkpoint epoch step avr_loss
-000002 2 264 0.0543
-000004 4 528 0.0492
-000006 6 792 0.0465
-000008 8 1056 0.0480
-000010 10 1320 0.0472
final 12 1584 0.0467

min avr_loss 0.0446 ยท final 0.0467

Inference (Krea 2 Turbo, Musubi Tuner)

Prepend the trigger woman to the prompt. Recommended: stack over the house style LoRA.

python src/musubi_tuner/krea2_generate_image.py \
    "woman, <scene>" \
    --dit turbo.safetensors --vae qwen_image_vae.safetensors \
    --text_encoder qwen3vl_4b_bf16.safetensors \
    --steps 8 --guidance_scale 1 --mu 1.15 --width 1024 --height 1280 \
    --attn_mode torch --lora_weight <checkpoint>.safetensors --lora_multiplier 1.0

Dataset samples

Representative image+caption pairs from the training set:

Millie-portrait-001-seedream.jpg

Millie-portrait-001-seedream

woman, close-up portrait facing the camera with her head slightly tilted, no hands visible, a soft closed-lip smile and a relaxed gaze, wearing a white scoop-neck top just visible at one shoulder, plain off-white background.

Millie-medium-001-seedream.jpg

Millie-medium-001-seedream

woman, medium shot, standing and facing the camera with arms relaxed at her sides and hands near her thighs, wearing a white long-sleeve tie-front crop top with a deep V neckline over a beige mini skirt with a side slit, a calm expression with lips slightly parted, plain light purple background.


Auto-generated by training watchdog.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Zaytron40k/millie-v4-krea2-lora

Base model

krea/Krea-2-Raw
Adapter
(933)
this model