SyFe LTX-2.3 ID-LoRA Checkpoints

Identity-driven audio-video LoRAs trained by SyFe on LTX-2.3 22B-dev. These checkpoints condition generation on a portrait and reference audio to preserve visual identity, vocal identity, and speaking motion in one pass.

Checkpoints

Run Training data Resolution / frames Rank Steps Status
id_lora_ours_768 3,376 talking clips from one show 768x320 / 121 128 3,000 Experimental; cloned voice and mouth motion validated
id_lora_ours_704 9,700 clips, 501 speakers, 20 shows 1280x704 / 121 128 6,000 Higher-resolution production candidate

Each run folder contains the final LoRA and its training configuration. The adapters use the ID-LoRA audio_ref_only_ic contract with negative-time reference-audio conditioning. They require an LTX-2.3 ID-LoRA-compatible pipeline; they are not standalone models.

Limitations

Lip synchronization is functional but not phoneme-perfect, and speech may truncate when the requested line does not fit the video duration. Identity quality depends strongly on portrait framing and face size.

Use is subject to the LTX-2 community license and the applicable rights for all reference media and generated identities.

Downloads last month
20
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for SyFeee/LTX-2.3-SyFe-ID-LoRA

Adapter
(108)
this model