RuWikiBench is a Russian benchmark dataset designed to evaluate analytical capabilities of large language models via Wikipedia-style article generation.
This repository contains the data used by the RuWikiBench pipeline: extracted source texts and link mappings.
The dataset is based on a manually selected subset of articles from the Russian internet encyclopedia "Ruwiki" (Рувики). The selection targets diverse topics with sufficient external… See the full description on the dataset page:
https://huggingface.co/datasets/NejimakiTori/RuWikiBench.