ByT5 multilingual P2G (tiny) โ€” 17M params

Phoneme-to-grapheme inversion of the harmonized retrain. Input: <lang>: phoneme tokens. Output: word spelling. Trained on a 4.12M-pair harmonized corpus.

Results (4k stratified test sample)

score
micro exact 0.397
macro exact 0.557

Files

  • HF-format weights at root (~70 MB)
  • onnx/ โ€” validated encoder+decoder pair

Licence

CC BY-SA 4.0

Downloads last month
39
Safetensors
Model size
17.9M params
Tensor type
F32
ยท
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for willwade/byt5-p2g-multilingual-tiny

Quantized
(5)
this model