Chanjing-Avatar 14B

Chanjing-Avatar 14B is an audio-driven 720p avatar video generation model based on Wan2.1-T2V-14B. It adds audio conditioning and LoRA adapters to the Wan video diffusion model.

Source code and complete inference instructions: chanjing-ai/Chanjing-Avatar

Chanjing-Avatar Model Family

The checkpoint contains audio modules, input projection, and LoRA adapters in BF16. The Wan2.1 base model and Wav2Vec audio encoder are required separately.

Chanjing-Avatar-14B/
|-- config.json
`-- diffusion_pytorch_model.safetensors
hf download cicada-ai/Chanjing-Avatar-14B \
  --local-dir models/Chanjing-Avatar-14B

Users are responsible for obtaining consent for source images and voices and for clearly disclosing synthetic media.

Downloads last month
4
Safetensors
Model size
0.6B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for cicada-ai/Chanjing-Avatar-14B

Finetuned
(77)
this model