This dataset is part of the benchmark introduced in the paper Graphusion: Leveraging Large Language Models for
Scientific Knowledge Graph Fusion and Construction in NLP Education. We also release more data in our GitHub page.
It contains 6 tasks designed for evaluating various aspects of reasoning, graph understanding, and language generation.
task1: Relation Judgment
task2: Prerequisite Prediction… See the full description on the dataset page:
https://huggingface.co/datasets/li-lab/tutorqa.