Dataset comprises 4 million+ logs of synthetic texts generated by large language models (LLMs) across 32 languages, leveraging 3 different GPT models for diverse, high-quality training data. Designed for text generation tasks, language model training, and NLP applications, supporting generative AI and text classification.- Get the data
Characteristic
Data
Description
Generated texts to achieve… See the full description on the dataset page:
https://huggingface.co/datasets/ud-nlp/LLM-Text-Generation-Dataset.