ltuncay commited on
Commit
e20d946
·
verified ·
1 Parent(s): b3c3ca4

Clarify Audio-JEPA training differences from the ICME checkpoint

Browse files
Files changed (2) hide show
  1. README.md +2 -0
  2. export_manifest.json +1 -1
README.md CHANGED
@@ -50,6 +50,8 @@ weight to the Speech, Music, and Environment category means.
50
 
51
  Audio-JEPA scores were supplied by the author for [run `jp6l70l6`](https://wandb.ai/tuncay-ludovic/audio%20embeddings/runs/jp6l70l6). The remaining scores are reported in the [project README](https://github.com/LudovicTuncay/audio-embeddings/blob/bb88bf790b1dcf8251c6b38e7a4766534adf33d3/README.md). These are reported research results, not a new benchmark run of the Transformers exports.
52
 
 
 
53
  ## Model and training
54
 
55
  | Property | Value |
 
50
 
51
  Audio-JEPA scores were supplied by the author for [run `jp6l70l6`](https://wandb.ai/tuncay-ludovic/audio%20embeddings/runs/jp6l70l6). The remaining scores are reported in the [project README](https://github.com/LudovicTuncay/audio-embeddings/blob/bb88bf790b1dcf8251c6b38e7a4766534adf33d3/README.md). These are reported research results, not a new benchmark run of the Transformers exports.
52
 
53
+ The Audio-JEPA row refers to the newer **16 kHz, 200,000-step** run, not the original ICME model (**32 kHz, 100,000 steps**). The author reports better results for this newer checkpoint.
54
+
55
  ## Model and training
56
 
57
  | Property | Value |
export_manifest.json CHANGED
@@ -29,7 +29,7 @@
29
  },
30
  "files": {
31
  "CODE_LICENSE": "8fe9e8b749cd4abedabcb3100df445db899e72192394d14cc2dbf24a40811af6",
32
- "README.md": "7039630b0e6bc883e37163ec3d2c383ae2265acd815cdc0b5761e1efc4c92c7a",
33
  "adapters.py": "ca0f26826763c0e48b7508da732242bb83adfeb4814e389ed3fde93ad648c728",
34
  "config.json": "fe8227548617416bd879efe1be3e9e3fc04270709250f09f01806780fe548789",
35
  "configuration_audio.py": "59f3a0b8db0df5af85e677ac33a4431595e9b2eff6d34c1e6a1778dd75a4568d",
 
29
  },
30
  "files": {
31
  "CODE_LICENSE": "8fe9e8b749cd4abedabcb3100df445db899e72192394d14cc2dbf24a40811af6",
32
+ "README.md": "662fd0b97f7dff054d662b4a0ff9b117b87c8a411275d4b715a4c19fdbd389a7",
33
  "adapters.py": "ca0f26826763c0e48b7508da732242bb83adfeb4814e389ed3fde93ad648c728",
34
  "config.json": "fe8227548617416bd879efe1be3e9e3fc04270709250f09f01806780fe548789",
35
  "configuration_audio.py": "59f3a0b8db0df5af85e677ac33a4431595e9b2eff6d34c1e6a1778dd75a4568d",