gemini-3.1-pro-hard-high-reasoning
- Type
- dataset
- Venue
- Roman1111111
- Year
- 2026
- Source
- huggingface
- Access
- free
- Language
- en
- Added
- 2026-07-17T20:00:45.865435+00:00
- Verified
- 2026-07-17T20:00:45.865435+00:00
Summary
This dataset represents the frontier of synthetic reasoning data, generated by **[Gemini 3.1 Pro](https://deepmind.google/technologies/gemini/)** (High Reasoning variant). While smaller in total token volume than its predecessors (5.6M tokens), this corpus prioritizes **logical density** and **multi-step verification**.
Keywords
hf-dataset qa · -language-modeling · -reasoning · -code json text datasets pandas polars mlcroissant code finance legal agent chemistry physics synthetic gemini-3.1-pro high-reasoning expert-level
Topics
QA, Language Modeling, Reasoning, Code
Research notes
- downloads=360; likes=58