GoLongRL
- Type
- dataset
- Venue
- Kwai-Klear
- Year
- 2026
- Source
- huggingface
- Access
- free
- Added
- 2026-07-17T20:02:48.880542+00:00
- Verified
- 2026-07-17T20:02:48.880542+00:00
Summary
This dataset is the RL training dataset for GoLongRL, targeting long-context capabilities of language models. It contains 23K training samples in total, with 9 types of reward functions.
Keywords
hf-dataset parquet text datasets dask polars mlcroissant has-paper
Research notes
- downloads=345; likes=23