Model type:
LLaVA-Next-Video is an open-source chatbot trained by fine-tuning LLM on multimodal instruction-following data.
Base LLM:
mistralai/Mistral-7B-Instruct-v0.2
Llama 2 is licensed under the LLAMA 2 Community License,
Copyright (c) Meta Platforms, Inc. All Rights Reserved.
A collection of 4 benchmarks, including 3 academic VQA benchmarks and 1 captioning benchmark.