Dataset automatically created during the evaluation run of model deepseek-chat
The dataset is composed of 18 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 49 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional configuration… See the full description on the dataset page:
https://huggingface.co/datasets/TheFinAI/lm-eval-xbrl-tagging-nen-0-shot-results-private.