Description: The TecDec dataset consists of Japanese texts, primarily policy-oriented or related to policy debates, labeled for technocratic versus deliberative frames. The dataset was created using a teacher-student paradigm with Human-in-the-Loop (HITL) active learning and includes both synthetic and human-generated samples.
Language: Japanese
Task Categories:
Text Classification: Binary classification (technocratic vs. deliberative).
Data Collection:
This… See the full description on the dataset page:
https://huggingface.co/datasets/ugo86/ja_tecdec_labels.