UltraData-SFT-2605
- Type
- dataset
- Venue
- OpenBMB (Hugging Face)
- Year
- 2026
- Source
- huggingface
- Access
- restricted
- Language
- English, Chinese
- Added
- 2026-07-17T20:18:03.821840+00:00
- Verified
- 2026-07-17T20:18:03.821840+00:00
Summary
UltraData-SFT-2605 is the full core-domain SFT dataset used in the post-training of MiniCPM5-1B-SFT, containing over 15 million Deep Thinking and Non-thinking training samples across math, code, knowledge, instruction following, and other domains. Every sample passes through a six-step High-Quality SFT Data Management Pipeline including query construction, answer quality filtering, single-data validation, and benchmark decontamination. The dual Deep Thinking / Non-thinking design trains models for both fast conversational responses and multi-step reasoning chains.
Keywords
sft instruction-tuning reasoning deep-thinking minicpm5 ultradata math code knowledge post-training
Topics
NLP / Instruction Tuning
Research notes
- Gated dataset (requires contact info). Part of the UltraData L0-L4 tiered data management framework. Released 2026.05.28 alongside MiniCPM5-1B.