This dataset was used to train the optillm-modernbert-large-bert
and optillm-bert-uncased router classifier model.
It was built by combining all instances of Arena Hard Auto
and MixEval datasets and running them through
the optillm proxy with gpt-4o-mini. We generated responses using optillm for all
the approaches and then evaluated them using LLM-as-Judge for Arena Hard Auto and ground truth for MixEval. These responses were
ranked and saved along with the number of tokens required for… See the full description on the dataset page:
https://huggingface.co/datasets/codelion/optillm-router-dataset.