Views
No views yet
| Dataset | #Images | Image Size | Spatial Resolution | #Total Captions |
|---|---|---|---|---|
| NWPU-Captions | 31,500 | 256 × 256 | ∼30-0.2m | 157,500 |
| RSICD | 10,921 | 224 × 224 | different resolutions | 54,605 |
| Sydney-Captions | 613 | 500 × 500 | 0.5m | 3,065 |
| UCM-Captions | 2,100 | 256 × 256 | ∼0.3m | 10,500 |
| Cap-4 | 45,134 | 224 × 224 | different resolutions | 225,670 |
1@article{silva2024large,
2 title={Large language models for captioning and retrieving remote sensing images},
3 author={Silva, Jo{\~a}o Daniel and Magalh{\~a}es, Jo{\~a}o and Tuia, Devis and Martins, Bruno},
4 journal={arXiv preprint arXiv:2402.06475},
5 year={2024}
6}