← Back to explorer

swallowmath2

Type
dataset
Venue
tokyotech-llm
Year
2026
Source
huggingface
Access
free
Language
en
Added
2026-07-17T19:59:21.418217+00:00
Verified
2026-07-17T19:59:21.418217+00:00

Summary

[SwallowMath-v2](https://huggingface.co/datasets/tokyotech-llm/swallow-math-v2) is a large-scale mathematical dataset containing **32 billion tokens**, developed as the successor to [SwallowMath-v1](https://huggingface.co/datasets/tokyotech-llm/swallow-math).

Keywords

hf-dataset language-modeling · -math json text datasets dask mlcroissant math has-paper

Topics

Language Modeling, Math

Research notes

  • downloads=11772; likes=32