The dataset contains 377 rows.
These paragraphs are extracted from authorized novels written by Khng Guan康原 and contain multiple sentences in Taiwanese Taigi written with Hanji.
The dataset maintains the original literary style and structure, making it useful for training language models, natural language processing (NLP), and Taiwanese literature research.
Number of rows: 377 (each representing a paragraph)
Features:
title: Book title… See the full description on the dataset page:
https://huggingface.co/datasets/IMA-Taiwan/taigi-literature-khg.