Video-Text-to-Text
Transformers
Safetensors
English
llava
text-generation
multi-modal
large-language-model
video-language-model
Instructions to use byminji/LLaVA-NeXT-13B-Video-FT with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use byminji/LLaVA-NeXT-13B-Video-FT with Transformers:
# Load model directly from transformers import AutoProcessor, AutoModelForSeq2SeqLM processor = AutoProcessor.from_pretrained("byminji/LLaVA-NeXT-13B-Video-FT") model = AutoModelForSeq2SeqLM.from_pretrained("byminji/LLaVA-NeXT-13B-Video-FT", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Download tokenizer.json from byminji/LLaVA-NeXT-13B-Video-FT: direct link, hf CLI and curl.
- Browser
- Download file 1.84 MB
-
https://huggingface.co/byminji/LLaVA-NeXT-13B-Video-FT/resolve/main/tokenizer.json
- Command line
-
hf download hf://byminji/LLaVA-NeXT-13B-Video-FT/tokenizer.json
-
curl -L -o tokenizer.json https://huggingface.co/byminji/LLaVA-NeXT-13B-Video-FT/resolve/main/tokenizer.json
1.84 MB
File too large to display, you can check the raw version instead.