This corpus was automatically generated by the Deku Corpus Builder for use in RAG-based AI applications.
text: The content text
source_url: Original source URL
source_title: Title of the source document
source_domain: Domain of the source
relevance_score: Relevance to the subject (0-1)
quality_score: Content quality score (0-1)
topics: JSON array of detected topics
character_count: Length of the text
subject_name:… See the full description on the dataset page:
https://huggingface.co/datasets/PhillyMac/FDR_Corpus.