Automatic Speech Recognition
Transformers
PyTorch
JAX
TensorBoard
Norwegian
whisper
audio
asr
hf-asr-leaderboard
Instructions to use NbAiLabArchive/scream_small_beta with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use NbAiLabArchive/scream_small_beta with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("automatic-speech-recognition", model="NbAiLabArchive/scream_small_beta")# Load model directly from transformers import AutoProcessor, AutoModelForSpeechSeq2Seq processor = AutoProcessor.from_pretrained("NbAiLabArchive/scream_small_beta") model = AutoModelForSpeechSeq2Seq.from_pretrained("NbAiLabArchive/scream_small_beta", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Download tokenizer_config.json from NbAiLabArchive/scream_small_beta: direct link, hf CLI and curl.
- Browser
- Download file 822 Bytes
-
https://huggingface.co/NbAiLabArchive/scream_small_beta/resolve/e89074af81ed5dbc9270f2eabc434b1d5dae00e9/tokenizer_config.json
- Command line
-
hf download hf://NbAiLabArchive/scream_small_beta@e89074af81ed5dbc9270f2eabc434b1d5dae00e9/tokenizer_config.json
-
curl -L -o tokenizer_config.json https://huggingface.co/NbAiLabArchive/scream_small_beta/resolve/e89074af81ed5dbc9270f2eabc434b1d5dae00e9/tokenizer_config.json
822 Bytes
| { | |
| "add_bos_token": false, | |
| "add_prefix_space": true, | |
| "bos_token": { | |
| "__type": "AddedToken", | |
| "content": "<|endoftext|>", | |
| "lstrip": false, | |
| "normalized": true, | |
| "rstrip": false, | |
| "single_word": false | |
| }, | |
| "clean_up_tokenization_spaces": true, | |
| "dropout": 0.1, | |
| "eos_token": { | |
| "__type": "AddedToken", | |
| "content": "<|endoftext|>", | |
| "lstrip": false, | |
| "normalized": true, | |
| "rstrip": false, | |
| "single_word": false | |
| }, | |
| "errors": "replace", | |
| "model_max_length": 1024, | |
| "pad_token": null, | |
| "processor_class": "WhisperProcessor", | |
| "return_attention_mask": false, | |
| "tokenizer_class": "WhisperTokenizer", | |
| "unk_token": { | |
| "__type": "AddedToken", | |
| "content": "<|endoftext|>", | |
| "lstrip": false, | |
| "normalized": true, | |
| "rstrip": false, | |
| "single_word": false | |
| } | |
| } | |