Video-Text-to-Text
Transformers
Safetensors
English
llava
text-generation
multimodal
Eval Results (legacy)
Instructions to use lmms-lab/LLaVA-Video-7B-Qwen2 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use lmms-lab/LLaVA-Video-7B-Qwen2 with Transformers:
# Load model directly from transformers import AutoProcessor, AutoModelForCausalLM processor = AutoProcessor.from_pretrained("lmms-lab/LLaVA-Video-7B-Qwen2") model = AutoModelForCausalLM.from_pretrained("lmms-lab/LLaVA-Video-7B-Qwen2", device_map="auto") - Notebooks
- Google Colab
- Kaggle
- Xet hash:
- 01d4fc2c6e5c5f1aee5b6b7b65839d4ecd33966cf73487d21cdec6c7c43245a3
- Size of remote file:
- 7.86 kB
- SHA256:
- afa15c93a4e0f20b7afc697641518f352df7f79c6138062bc4f161035687b85a
·
Xet efficiently stores Large Files inside Git, intelligently splitting files into unique chunks and accelerating uploads and downloads. More info.