Dataset contains a large (3.5M+ sentence) knowledge base of generic sentences. This is the first large resource to contain naturally occurring generic sentences, rich in high-quality, general, semantically complete statements. All GenericsKB sentences are annotated with their topical term, surrounding context (sentences), and a (learned) confidence. We also release GenericsKB-Best (1M+ sentences), containing the best-quality… See the full description on the dataset page: https://huggingface.co/datasets/community-datasets/generics_kb.