This repository contains the outcomes of your submitted models that have been evaluated through the Open LLM Leaderboard. Our goal is to shed light on the cutting-edge Large Language Models (LLMs) and chatbots, enabling you to make well-informed decisions regarding your chosen application.
The evaluation process involves running your models against several benchmarks from the Eleuther AI Harness, a unified framework for… See the full description on the dataset page:
https://huggingface.co/datasets/open-llm-leaderboard-old/results.