ForestFireVLM-3B is a specialized vision-language model fine-tuned from
Qwen2.5-VL-3B-Instruct specifically for forest fire detection and analysis tasks. This model is designed to identify and analyze various aspects of forest fires from aerial imagery, including smoke detection, flame visibility, fire characteristics, and potential hazards.
The
LLaMA-Factory framework was used for fine-tuning this model.
Evaluations were done with our code available on
GitHub, using the
ForestFireInsights-Eval dataset.
This model is associated with research currently under peer review with MDPI. Please cite our paper when using this model: