Audio-Text-to-Text
Transformers
Safetensors
English
audioflamingo3
text2text-generation
audio
reasoning
audio understanding
ASR
Instructions to use nvidia/audio-flamingo-3-hf with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use nvidia/audio-flamingo-3-hf with Transformers:
# pip install -U transformers accelerate # Load model directly from transformers import AutoProcessor, AutoModelForSeq2SeqLM processor = AutoProcessor.from_pretrained("nvidia/audio-flamingo-3-hf") model = AutoModelForSeq2SeqLM.from_pretrained("nvidia/audio-flamingo-3-hf", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Download added_tokens.json from nvidia/audio-flamingo-3-hf: direct link, hf CLI and curl.
- Browser
- Download file 762 Bytes
-
https://huggingface.co/nvidia/audio-flamingo-3-hf/resolve/main/added_tokens.json
- Command line
-
hf download hf://nvidia/audio-flamingo-3-hf/added_tokens.json
-
curl -L -o added_tokens.json https://huggingface.co/nvidia/audio-flamingo-3-hf/resolve/main/added_tokens.json
762 Bytes
| { | |
| "</tool_call>": 151658, | |
| "<image>": 151666, | |
| "<sound>": 151669, | |
| "<speech>": 151668, | |
| "<tool_call>": 151657, | |
| "<vila/sentinel>": 151665, | |
| "<vila/video>": 151667, | |
| "<|box_end|>": 151649, | |
| "<|box_start|>": 151648, | |
| "<|endoftext|>": 151643, | |
| "<|file_sep|>": 151664, | |
| "<|fim_middle|>": 151660, | |
| "<|fim_pad|>": 151662, | |
| "<|fim_prefix|>": 151659, | |
| "<|fim_suffix|>": 151661, | |
| "<|im_end|>": 151645, | |
| "<|im_start|>": 151644, | |
| "<|image_pad|>": 151655, | |
| "<|object_ref_end|>": 151647, | |
| "<|object_ref_start|>": 151646, | |
| "<|quad_end|>": 151651, | |
| "<|quad_start|>": 151650, | |
| "<|repo_name|>": 151663, | |
| "<|video_pad|>": 151656, | |
| "<|vision_end|>": 151653, | |
| "<|vision_pad|>": 151654, | |
| "<|vision_start|>": 151652, | |
| "[BOS]": 151670, | |
| "[PAD]": 151671 | |
| } | |