MuQ-MuLan-large β€” ONNX

An ONNX export of OpenMuQ/MuQ-MuLan-large (revision 2e01c79), which embeds music and text in the same 512-dimensional space.

  • audio.onnx β€” mono waveform at 24 kHz (float32, 10 s windows of 240 000 samples) β†’ one L2-normalised 512-d vector per window. The mel front-end is inside the graph.
  • text.onnx + tokenizer.json β€” text (up to 64 tokens) β†’ an L2-normalised 512-d vector.
  • manifest.json β€” inputs, outputs and file hashes.

Outputs match the original PyTorch model (cosine similarity β‰₯ 0.99999). Converted, not retrained.

License

CC BY-NC 4.0, like the original weights: non-commercial use only.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for YvanM/MuQ-MuLan-large-onnx

Quantized
(2)
this model