| Paper | Github | Homepage |
We introduce SpreadsheetBench, a challenging spreadsheet manipulation benchmark exclusively derived from real-world scenarios, designed
to immerse current large language models (LLMs) in the actual workflow of spreadsheet users. Unlike existing benchmarks that rely on
synthesized queries and simplified spreadsheet files, SpreadsheetBench is built from 912 real questions gathered… See the full description on the dataset page:
https://huggingface.co/datasets/KAKA22/SpreadsheetBench.