This model card was not created by the original authors of the model.
Instead, it represents a conversion of the original model to the Hugging Face format.
The original model can be found
here.
This conversion was performed using the Transformers conversion tutorial.
The purpose of this conversion was to make the model compatible with vllm serving, as the original model is currently not supported by
vllm.
This model is a conversion of the original LLava Next Video weights. However, a better-converted version of this model has been made available by the VLLM contributors (
here).
Made of original LLava One Vision weights.
Since this is a converted model, it may exhibit bugs or produce outputs that are inconsistent with the original model's performance and accuracy.
These discrepancies are likely due to differences in compatibility and are not reflective of the original model's quality.
This model is best used for experimentation and tasks compatible with
vllm serving.
It is recommended to cross-reference outputs with the original model for critical applications to ensure reliability and correctness.
We acknowledge the original creators of the model for their work and contributions.
The conversion and hosting of this model aim to expand its usability while maintaining the integrity of the
original work.
This model is provided "as is" without warranty of any kind.
Any issues, bugs, or errors encountered are solely related to the conversion process or platform compatibility and are not attributable to the
original model.