This is just a stub from the original repo
All credits go to the original author David Roher.
A structured, comprehensive, and multilingual etymology dataset created by parsing Wiktionary's etymology sections. Key features:
4.2+ million etymological relationships between 2.0+ million terms in 3300+ languages/dialects
31 different types of etymological relations, distinguishing between inheritance, borrowing, etc.
Hierarchical data that preserves relationship… See the full description on the dataset page:
https://huggingface.co/datasets/Nickmancol/etymology.