π DMind Benchmark has been accepted to the Dataset & Benchmark Track of KDD 2026! ππ
A comprehensive framework for evaluating large language models (LLMs) on blockchain, cryptocurrency, and Web3 knowledge across multiple domains.
| Paper | Dataset |
Overall performance of all evaluated LLMs on the DMind Benchmark
This project provides tools to benchmark AI models on their understanding of blockchain concepts through both⦠See the full description on the dataset page:
https://huggingface.co/datasets/DMindAI/DMind_Benchmark.