π En-Bn-Code-Mixed-Two-Class-Sentiment-Dataset
The En-Bn-Code-Mixed-Two-Class-Sentiment-Dataset is a multilingual dataset of 100,000 product review texts designed for code-mixed sentiment analysis involving English, Bengali, and Roman Bengali.
Each record includes:
π Id
π ProductId
π¬ Code-Mixed-Text
π‘ Sentiment
The dataset captures diverse linguistic styles, authentic code-mixing, and real-world sentiment patterns from multilingual digital communication.
π Text Distribution
The dataset⦠See the full description on the dataset page:
https://huggingface.co/datasets/DaliaBarua/En-Bn-Code-Mixed-Two-Class-Sentiment-Dataset.