Audio-Text-to-Text
Transformers
English
Chinese
transformer
multimodal
vqa
text
audio
Eval Results (legacy)
Instructions to use zeroMN/SHMT with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use zeroMN/SHMT with Transformers:
# Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("zeroMN/SHMT", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Download vision_encoder/tokenizer_config.json from zeroMN/SHMT: direct link, hf CLI and curl.
- Browser
- Download file 767 Bytes
-
https://huggingface.co/zeroMN/SHMT/resolve/main/vision_encoder/tokenizer_config.json
- Command line
-
hf download hf://zeroMN/SHMT/vision_encoder/tokenizer_config.json
-
curl -L -o tokenizer_config.json https://huggingface.co/zeroMN/SHMT/resolve/main/vision_encoder/tokenizer_config.json
767 Bytes
| { | |
| "add_prefix_space": false, | |
| "added_tokens_decoder": { | |
| "49406": { | |
| "content": "<|startoftext|>", | |
| "lstrip": false, | |
| "normalized": true, | |
| "rstrip": false, | |
| "single_word": false, | |
| "special": true | |
| }, | |
| "49407": { | |
| "content": "<|endoftext|>", | |
| "lstrip": false, | |
| "normalized": false, | |
| "rstrip": false, | |
| "single_word": false, | |
| "special": true | |
| } | |
| }, | |
| "bos_token": "<|startoftext|>", | |
| "clean_up_tokenization_spaces": false, | |
| "do_lower_case": true, | |
| "eos_token": "<|endoftext|>", | |
| "errors": "replace", | |
| "extra_special_tokens": {}, | |
| "model_max_length": 77, | |
| "pad_token": "<|endoftext|>", | |
| "tokenizer_class": "CLIPTokenizer", | |
| "unk_token": "<|endoftext|>" | |
| } | |