Dataset Card for Multilingual Synthetic Online Conversations
Dataset Summary
This dataset contains multilingual translations of the Synthetic Online Conversations (SOC-2508) dataset. Each conversation from the original dataset has been translated into French, Italian, German, Spanish, providing over 1,180 synthetically generated, multi-turn online conversations in multiple languages.
The translations were generated using google/gemma-3n-E4B-it with vLLM as the inference… See the full description on the dataset page: https://huggingface.co/datasets/marcodsn/SOC-2508-MULTI.