Mirror nv-modelcard++/explainability.md from nvidia/bigvgan_v2_44khz_128band_512x@95a9d1dc

Browse files

Files changed (1) hide show

encoders/nvidia/bigvgan_v2_44khz_128band_512x/nv-modelcard++/explainability.md +13 -0

encoders/nvidia/bigvgan_v2_44khz_128band_512x/nv-modelcard++/explainability.md ADDED Viewed

	@@ -0,0 +1,13 @@

+| Field                                                                                                 | Response                                                                                                                                                                                                               |
+| :---------------------------------------------------------------------------------------------------- | :--------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
+| Intended Application & Domain:                                                                        | Generating waveform from mel spectrogram.                                                                                                                                                                              |
+| Model Type:                                                                                           | Convolutional Neural Network (CNN)                                                                                                                                                                                     |
+| Intended Users:                                                                                       | This model is intended for developers to synthesize and generate waveforms from the AI-generated mel spectrograms.                                                                                                     |
+| Output:                                                                                               | Audio Waveform                                                                                                                                                                                                         |
+| Describe how the model works:                                                                         | Model generates audio waveform corresponding to the input mel spectrogram.                                                                                                                                             |
+| Name the adversely impacted groups this has been tested to deliver comparable outcomes regardless of: | Not Applicable                                                                                                                                                                                                         |
+| Technical Limitations:                                                                                | This may not perform well on synthetically-generated mel spectrograms that deviate significantly from the profile of mel spectrograms on which this was trained.                                                       |
+| Verified to have met prescribed NVIDIA quality standards:                                             | Yes                                                                                                                                                                                                                    |
+| Performance Metrics:                                                                                  | Perceptual Evaluation of Speech Quality (PESQ), Virtual Speech Quality Objective Listener (VISQOL), Multi-resolution STFT (MRSTFT), Mel cepstral distortion (MCD), Periodicity RMSE, Voice/Unvoiced F1 Score (V/UV F1) |
+| Potential Known Risks:                                                                                | This model may generate low-quality or distorted soundwaves.                                                                                                                                                           |
+| Licensing:                                                                                            | https://github.com/NVIDIA/BigVGAN/blob/main/LICENSE                                                                                                                                                                    |