A curated multi-domain dataset for powering the LexiMind HuggingFace Space demo. Contains 1,219 items spanning academic papers, literary works, social media text, and curated technical blog posts — each annotated with topic and emotion labels.
No news articles. The LexiMind model is trained on ArXiv papers and Project Gutenberg books; news data produced poor summarization results due to domain mismatch.