Feature Extraction
Transformers
Safetensors
audio_embeddings
audio
custom_code
self-supervised-learning
audio-embeddings
best-rq-2
audioset
Instructions to use ltuncay/BEST-RQ-ViT with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use ltuncay/BEST-RQ-ViT with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("feature-extraction", model="ltuncay/BEST-RQ-ViT", trust_remote_code=True)# Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("ltuncay/BEST-RQ-ViT", trust_remote_code=True, device_map="auto") - Notebooks
- Google Colab
- Kaggle
Clarify Audio-JEPA training differences from the ICME checkpoint
Browse files- README.md +2 -0
- export_manifest.json +1 -1
README.md
CHANGED
|
@@ -50,6 +50,8 @@ weight to the Speech, Music, and Environment category means.
|
|
| 50 |
|
| 51 |
Audio-JEPA scores were supplied by the author for [run `jp6l70l6`](https://wandb.ai/tuncay-ludovic/audio%20embeddings/runs/jp6l70l6). The remaining scores are reported in the [project README](https://github.com/LudovicTuncay/audio-embeddings/blob/bb88bf790b1dcf8251c6b38e7a4766534adf33d3/README.md). These are reported research results, not a new benchmark run of the Transformers exports.
|
| 52 |
|
|
|
|
|
|
|
| 53 |
## Model and training
|
| 54 |
|
| 55 |
| Property | Value |
|
|
|
|
| 50 |
|
| 51 |
Audio-JEPA scores were supplied by the author for [run `jp6l70l6`](https://wandb.ai/tuncay-ludovic/audio%20embeddings/runs/jp6l70l6). The remaining scores are reported in the [project README](https://github.com/LudovicTuncay/audio-embeddings/blob/bb88bf790b1dcf8251c6b38e7a4766534adf33d3/README.md). These are reported research results, not a new benchmark run of the Transformers exports.
|
| 52 |
|
| 53 |
+
The Audio-JEPA row refers to the newer **16 kHz, 200,000-step** run, not the original ICME model (**32 kHz, 100,000 steps**). The author reports better results for this newer checkpoint.
|
| 54 |
+
|
| 55 |
## Model and training
|
| 56 |
|
| 57 |
| Property | Value |
|
export_manifest.json
CHANGED
|
@@ -29,7 +29,7 @@
|
|
| 29 |
},
|
| 30 |
"files": {
|
| 31 |
"CODE_LICENSE": "8fe9e8b749cd4abedabcb3100df445db899e72192394d14cc2dbf24a40811af6",
|
| 32 |
-
"README.md": "
|
| 33 |
"adapters.py": "ca0f26826763c0e48b7508da732242bb83adfeb4814e389ed3fde93ad648c728",
|
| 34 |
"config.json": "fe8227548617416bd879efe1be3e9e3fc04270709250f09f01806780fe548789",
|
| 35 |
"configuration_audio.py": "59f3a0b8db0df5af85e677ac33a4431595e9b2eff6d34c1e6a1778dd75a4568d",
|
|
|
|
| 29 |
},
|
| 30 |
"files": {
|
| 31 |
"CODE_LICENSE": "8fe9e8b749cd4abedabcb3100df445db899e72192394d14cc2dbf24a40811af6",
|
| 32 |
+
"README.md": "662fd0b97f7dff054d662b4a0ff9b117b87c8a411275d4b715a4c19fdbd389a7",
|
| 33 |
"adapters.py": "ca0f26826763c0e48b7508da732242bb83adfeb4814e389ed3fde93ad648c728",
|
| 34 |
"config.json": "fe8227548617416bd879efe1be3e9e3fc04270709250f09f01806780fe548789",
|
| 35 |
"configuration_audio.py": "59f3a0b8db0df5af85e677ac33a4431595e9b2eff6d34c1e6a1778dd75a4568d",
|