← Back to explorer

GoLongRL

Type
dataset
Venue
Kwai-Klear
Year
2026
Source
huggingface
Access
free
Added
2026-07-17T20:02:48.880542+00:00
Verified
2026-07-17T20:02:48.880542+00:00

Summary

This dataset is the RL training dataset for GoLongRL, targeting long-context capabilities of language models. It contains 23K training samples in total, with 9 types of reward functions.

Keywords

hf-dataset parquet text datasets dask polars mlcroissant has-paper

Research notes

  • downloads=345; likes=23