Video-Text-to-Text
Transformers
Safetensors
English
audiovisualflamingo
text-generation
audio
video
image
audio-visual
multimodal
reasoning
video understanding
long-video
audio understanding
ASR
timestamp-grounding
instruction-tuned
dynamic-s2
CRTE
TAVIT
Instructions to use nvidia/nemotron-labs-audio-visual-flamingo-hf with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use nvidia/nemotron-labs-audio-visual-flamingo-hf with Transformers:
# Load model directly from transformers import AutoModelForSeq2SeqLM model = AutoModelForSeq2SeqLM.from_pretrained("nvidia/nemotron-labs-audio-visual-flamingo-hf", device_map="auto") - Notebooks
- Google Colab
- Kaggle

- Xet hash:
- a91e73298f8cd38275936d9252d600a0171e795ec0ee491d33df3966092b840a
- Size of remote file:
- 193 kB
- SHA256:
- 0446ca36643f6bc5b17653bc1ab9b18c895c159ab1db9e927d20cb8813888a8c
·
Xet efficiently stores Large Files inside Git, intelligently splitting files into unique chunks and accelerating uploads and downloads. More info.