Automatic Speech Recognition
Transformers
Safetensors
fun_asr_nano
text-generation
speech-recognition
asr
end-to-end
multilingual
streaming
arxiv:2407.04051
Instructions to use FunAudioLLM/Fun-ASR-Nano-2512-hf with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use FunAudioLLM/Fun-ASR-Nano-2512-hf with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("automatic-speech-recognition", model="FunAudioLLM/Fun-ASR-Nano-2512-hf")# Load model directly from transformers import AutoModelForSeq2SeqLM model = AutoModelForSeq2SeqLM.from_pretrained("FunAudioLLM/Fun-ASR-Nano-2512-hf", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Regenerate checkpoint for current Transformers layout
Browse filesRegenerated from the original Fun-ASR-Nano ModelScope checkpoint after the Whisper component refactor. Local conversion loaded 1,540 tensors with zero missing, unexpected, or mismatched keys; H100 Chinese, English, default-prompt, and batch inference matched expected outputs.
- tokenizer_config.json +1 -1
tokenizer_config.json
CHANGED
|
@@ -21,7 +21,7 @@
|
|
| 21 |
"<|video_pad|>"
|
| 22 |
],
|
| 23 |
"is_local": true,
|
| 24 |
-
"local_files_only":
|
| 25 |
"model_max_length": 131072,
|
| 26 |
"pad_token": "<|endoftext|>",
|
| 27 |
"processor_class": "FunAsrNanoProcessor",
|
|
|
|
| 21 |
"<|video_pad|>"
|
| 22 |
],
|
| 23 |
"is_local": true,
|
| 24 |
+
"local_files_only": false,
|
| 25 |
"model_max_length": 131072,
|
| 26 |
"pad_token": "<|endoftext|>",
|
| 27 |
"processor_class": "FunAsrNanoProcessor",
|