hanwenggggggg's picture
Document Community-1 provenance and license scope
df2625a
|
Raw History Blame Contribute Delete
1.93 kB
# Attribution and license scope
This is a mixed-license-scope repository. Fluid Inference intends the included
CC-BY-4.0 license text to cover only the following exact Community-1-derived
artifacts and their uncompiled counterparts listed in `provenance.json`:
- `Segmentation.mlmodelc`
- `FBank.mlmodelc`
- `Embedding.mlmodelc`
- `PLDA.mlmodelc`
- `PldaRho.mlmodelc`
- `plda-parameters.json`
- `xvector-transform.json`
The underlying Community-1 pipeline is published by pyannote under CC-BY-4.0:
https://huggingface.co/pyannote/speaker-diarization-community-1
The PLDA parameters originate from Brno University of Technology / BUT
Speech@FIT. The rights holder explicitly licenses `plda.npz` and
`xvec_transform.npz` under CC-BY-4.0, including commercial use:
https://huggingface.co/BUT-FIT/diarizen-wavlm-large-s80-md/blob/6285693ddd5b38e8229acb93f864f3d04a82bee1/plda/LICENSE
Fluid Inference converted the PyTorch components to Core ML, introduced fixed
and enumerated input shapes, applied mixed-precision storage where recorded in
the model metadata, separated the FBank frontend from the embedding backend,
and compiled packages for Apple platforms.
Appropriate attribution should identify pyannote, WeSpeaker, BUT Speech@FIT,
and Fluid Inference, retain the citations in README.md, link CC-BY-4.0, and
indicate that the files are modified Core ML conversions.
## Legacy exclusions
`pyannote_segmentation.mlmodelc`, `wespeaker.mlmodelc`,
`wespeaker_v2.mlmodelc`, `wespeaker_int8.mlmodelc`, and their packages predate
the supported Community-1 artifact set. They are retained for compatibility but
are not covered by this Community-1 provenance and license-scope confirmation.
Their original source and licensing must be evaluated separately.
Fluid Inference can grant rights only to the extent it is authorized to do so.
This notice does not restrict or replace rights granted directly by upstream
licensors.