← Back to explorer

gemini-3.1-pro-hard-high-reasoning

Type
dataset
Venue
Roman1111111
Year
2026
Source
huggingface
Access
free
Language
en
Added
2026-07-17T20:00:45.865435+00:00
Verified
2026-07-17T20:00:45.865435+00:00

Summary

This dataset represents the frontier of synthetic reasoning data, generated by **[Gemini 3.1 Pro](https://deepmind.google/technologies/gemini/)** (High Reasoning variant). While smaller in total token volume than its predecessors (5.6M tokens), this corpus prioritizes **logical density** and **multi-step verification**.

Keywords

hf-dataset qa · -language-modeling · -reasoning · -code json text datasets pandas polars mlcroissant code finance legal agent chemistry physics synthetic gemini-3.1-pro high-reasoning expert-level

Topics

QA, Language Modeling, Reasoning, Code

Research notes

  • downloads=360; likes=58