SmallEval is a curated collection of lightweight evaluation datasets specifically designed for testing Large Language Models (LLMs) in browser environments. Each dataset is carefully subsampled to maintain a small footprint while preserving the evaluation quality.
The primary goal of SmallEval is to enable efficient evaluation of LLMs directly in web browsers. Traditional evaluation datasets are often⦠See the full description on the dataset page:
https://huggingface.co/datasets/smalleval/mmlu-nano.