Word2Bezbar: Word2Vec Models for French Rap Lyrics
Overview
Word2Bezbar are Word2Vec models trained on french rap lyrics sourced from Genius. Tokenization has been done using NLTK french word_tokenze function, with a prior processing to remove french oral contractions. Used dataset size was 323MB, corresponding to 77M tokens.
The model captures the semantic relationships between words in the context of french rap, providing a useful tool for studies associated to french slang and music lyrics analysis.
Model Details
Size of this model is medium
Parameter
Value
Dimensionality
200
Window Size
10
Epochs
20
Algorithm
CBOW
Versions
This model has been trained with the followed software versions
This model is designed for academic and research purposes only. It is not intended for commercial use. The creators of this model do not endorse or promote any specific views or opinions that may be represented in the dataset.
Please mention @RapMinerz if you use our models
Contact
For any questions or issues, please contact the repository owner, RapMinerz, at rapminerz.contact@gmail.com.