Synthetic Russian + English prompt → answer dataset for SFT/chat fine-tuning experiments.
Current local build: 20,000 examples.
greetings and warm opening conversations;
more small talk and human-like supportive replies;
more coding categories;
debugging prompts;
code review prompts;
tests and refactoring prompts;
API design;
data analysis;… See the full description on the dataset page:
https://huggingface.co/datasets/SonexaAI/ru_eng-humanity-dataset.