Instructions to use Avdpro/FlashHead-Lite-MLX with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use Avdpro/FlashHead-Lite-MLX with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir FlashHead-Lite-MLX Avdpro/FlashHead-Lite-MLX
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
FlashHead Lite MLX
Self-contained checkpoint for AI2Apps FlashHead native MLX Model Worker. Runtime and model Python code are distributed separately.
Includes Lite DiT, its VAE and Wav2Vec2 base. Original tensor layouts are converted to MLX on load; Pro Wan VAE was converted offline from a pure tensor state dictionary to safetensors. No Torch is required for inference.
Fixed upstream revisions, exact runtime file list, sizes and SHA-256 values are in ai2apps-checkpoint.json.
Sources: SoulX-FlashHead, official weights, Wav2Vec2. Upstream model cards declare Apache-2.0. Preserve LICENSE and NOTICE.md.
512x512, 25 FPS, four denoising steps. This is a local offline generation path; no real-time streaming claim.
Hardware compatibility
Log In to add your hardware
Quantized