akouo — speaker diarization (Core ML)

Core ML speaker segmentation and embedding models, used by akouo for iOS to work out who spoke when — entirely on device.

This is a mirror, hosted so the app does not depend on a third-party repository at runtime. The weights are unmodified.

Provenance

Licence — attribution required

CC-BY-4.0.

The upstream pyannote models carry mixed terms: pyannote/segmentation-3.0 is MIT, while pyannote/wespeaker-voxceleb-resnet34-LM and pyannote/speaker-diarization-community-1 are CC-BY-4.0. Where bundled components differ, the most restrictive governs — so CC-BY-4.0 applies to this repository as a whole.

If you redistribute these weights or ship them inside an application, you must credit pyannote visibly. This is a condition of the licence, not a courtesy.

Argmax does not declare a licence on the source repository. The licence stated here is a good-faith reading of what these weights derive from, not a grant by tinypocket.

Downloads last month
-
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support