Video-Text-to-Text
Transformers
Safetensors
English
qwen2
text-generation
Action
Video
MQA
multimodal
VLM
LLaVAction
MLLMs
Eval Results (legacy)
text-generation-inference
Instructions to use MLAdaptiveIntelligence/LLaVAction-7B with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use MLAdaptiveIntelligence/LLaVAction-7B with Transformers:
# Load model directly from transformers import AutoTokenizer, AutoModelForCausalLM tokenizer = AutoTokenizer.from_pretrained("MLAdaptiveIntelligence/LLaVAction-7B") model = AutoModelForCausalLM.from_pretrained("MLAdaptiveIntelligence/LLaVAction-7B", device_map="auto") - Notebooks
- Google Colab
- Kaggle
| { | |
| "<image>": 151646, | |
| "<|endoftext|>": 151643, | |
| "<|im_end|>": 151645, | |
| "<|im_start|>": 151644 | |
| } | |