swallowmath2
- Type
- dataset
- Venue
- tokyotech-llm
- Year
- 2026
- Source
- huggingface
- Access
- free
- Language
- en
- Added
- 2026-07-17T19:59:21.418217+00:00
- Verified
- 2026-07-17T19:59:21.418217+00:00
Summary
[SwallowMath-v2](https://huggingface.co/datasets/tokyotech-llm/swallow-math-v2) is a large-scale mathematical dataset containing **32 billion tokens**, developed as the successor to [SwallowMath-v1](https://huggingface.co/datasets/tokyotech-llm/swallow-math).
Keywords
hf-dataset language-modeling · -math json text datasets dask mlcroissant math has-paper
Topics
Language Modeling, Math
Research notes
- downloads=11772; likes=32