This is a fork of
liuhaotian/llava-v1.6-mistral-7b to be fully
compatible for inference with
SGLang.
No other changes were made.
Model type:
LLaVA is an open-source chatbot trained by fine-tuning LLM on multimodal instruction-following data.
It is an auto-regressive language model, based on the transformer architecture.
Base LLM:
mistralai/Mistral-7B-Instruct-v0.2
A collection of 12 benchmarks, including 5 academic VQA benchmarks and 7 recent benchmarks specifically proposed for instruction-following LMMs.