IT-RAG-Bench is a synthetic Italian-language retrieval benchmark designed to evaluate dense embedding models on document retrieval and Retrieval-Augmented Generation (RAG) tasks in Italian.
This dataset is the companion resource for the paper:
Benchmarking Google Embeddings 2 against Open-Source Models for Multilingual Dense Retrieval and RAG SystemsStefano Cirillo, Domenico Desiato, Giuseppe Polese, Giandomenico Solimando — arXiv… See the full description on the dataset page:
https://huggingface.co/datasets/DAISLab-Unisa/it-rag-bench.