Blossom is a powerful open-source conversational large language model that provides reproducible post-training data, dedicated to delivering an open, powerful, and cost-effective locally accessible general-purpose model for everyone.
You can find the training data here:
Blossom-V6.2-SFT-Stage1 (1 epoch)、
Blossom-V6.2-SFT-Stage2 (3 epoch).
Primarily employs three cost-effective models: Deepseek-V3.1, Gemini 2.5 Flash, and Qwen3-235B-A22B-Instruct-2507 (denoted as A, B, C)—to regenerate responses under different scenarios using tailored synthesis strategies.
Further technical details will be released in the future. The data is synthesized by the
🌸BlossomData framework.