krea2-style-v5

Krea 2 style LoRA (digital art). Trained on Krea 2 RAW, intended for inference on Krea 2 Turbo.

Status: ✅ training complete

Training configuration

Field Value
Base model (train) krea/Krea-2-Raw (raw.safetensors, 12B DiT, undistilled)
Inference target krea/Krea-2-Turbo (8-step distilled) — RAW-train / Turbo-infer
VAE Qwen-Image VAE (qwen_image_vae.safetensors)
Text encoder Qwen3-VL-4B-Instruct (qwen3vl_4b_bf16.safetensors)
Trainer Musubi Tuner (kohya-ss), networks.lora_krea2
LoRA rank/dim 32, alpha 32, all 264 Linear layers of the DiT
Optimizer adamw8bit, lr 1e-4
Precision bf16, gradient checkpointing, SDPA
Timestep sampling krea2_shift (resolution-aware), weighting_scheme none
Dataset 427 images + paired captions (221 style/SFW + 9 male body-type + 13 nude/flaccid + 67 Genatomy anatomy close-ups + 117 v5 add-on: seedream mature/young men & women, character shots), batch 4, 1024 multi-res buckets, num_repeats 1
Schedule 40 epochs = 4880 steps, save every 5 epochs, seed 42
Hardware 1x NVIDIA H100 80GB (RunPod)

Loss per checkpoint

avr_loss is a rolling average — flow-matching loss is near-flat, so use it as a sanity check and pick the best checkpoint visually.

checkpoint epoch step avr_loss
-000005 5 610 0.0664
-000010 10 1220 0.0664
-000015 15 1830 0.0680
-000020 20 2440 0.0671
-000025 25 3050 0.0648
-000030 30 3660 0.0630
-000035 35 4270 0.0629
final 40 4880 0.0633

min avr_loss 0.0509 · final 0.0633

Inference (Krea 2 Turbo, Musubi Tuner)

python src/musubi_tuner/krea2_generate_image.py \
    "your prompt" \
    --dit turbo.safetensors --vae qwen_image_vae.safetensors \
    --text_encoder qwen3vl_4b_bf16.safetensors \
    --steps 8 --guidance_scale 1 --mu 1.15 --width 1024 --height 1024 \
    --attn_mode torch --lora_weight <checkpoint>.safetensors --lora_multiplier 1.0

Dataset samples

Two representative image+caption pairs from the 427-image training set:

mat-amb_m09.png

mat-amb_m09

A heavyset middle-aged man with light skin, a receding hairline and short brown hair. He lifts his chin and looks upward, brows slightly lowered, with a satisfied closed-lip smile; his expression is proud. He rests his fist under his chin, the other arm folded across his belly, wearing a blue denim jacket over a white T-shirt, a brown belt and blue jeans. Three-quarter length, side profile at eye level. Plain wine red background.

young_yw04.png

young_yw04

A young adult woman with subtle South Asian features, brown skin, wavy shoulder-length black hair with lighter ends and gold hoop earrings. She raises her brows and beams with an open smile showing teeth, gazing up to the side; her expression is cheerful. She holds a takeaway coffee cup, her other hand in her pocket, wearing an olive green parka with pink lining over a grey turtleneck and jeans. Hip-up, three-quarter view from a low angle. Plain dark plum background.


Auto-generated by training watchdog.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Zaytron40k/krea2-style-v5

Base model

krea/Krea-2-Raw
Adapter
(933)
this model