Irodori-TTS model assets for audio.cpp (NovelViewer)

Runtime model assets for the endo5501/audio.cpp fork (Irodori-TTS engine used by NovelViewer). This repository repackages the minimal file set required by the audio.cpp irodori_tts safetensors loader, laid out as sibling directories:

Irodori-TTS-600M-v3-VoiceDesign/
  model.safetensors
  model_config.json
llm-jp-3-150m/
  tokenizer.json
Semantic-DACVAE-Japanese-32dim/
  weights.safetensors   (converted from upstream weights.pth)

Sources and licenses

Asset Upstream License
Irodori-TTS-600M-v3-VoiceDesign Aratako/Irodori-TTS-600M-v3-VoiceDesign MIT (+ ethical restrictions, see below)
llm-jp-3-150m tokenizer llm-jp/llm-jp-3-150m Apache-2.0
Semantic-DACVAE-Japanese-32dim Aratako/Semantic-DACVAE-Japanese-32dim MIT

Semantic-DACVAE-Japanese-32dim/weights.safetensors is a format conversion (PyTorch weights.pth โ†’ safetensors) of the upstream checkpoint; weights are unmodified.

Ethical restrictions (inherited from Irodori-TTS)

In addition to the MIT license terms, the upstream Irodori-TTS model states ethical restrictions on use (e.g., prohibiting impersonation without consent and unlawful use). See the upstream model card for the authoritative text: https://huggingface.co/Aratako/Irodori-TTS-600M-v3-VoiceDesign

Downloads last month
31
GGUF
Hardware compatibility
Log In to add your hardware

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support