ImageNet (ILSVRC 2012)
- Type
- dataset
- Venue
- Princeton University / Stanford University (ILSVRC)
- Year
- 2026
- Source
- huggingface
- Access
- restricted
- Language
- English
- Added
- 2026-07-17T20:18:03.814427+00:00
- Verified
- 2026-07-17T20:18:03.814427+00:00
Summary
ImageNet ILSVRC 2012 is the most widely used image classification benchmark, containing 1,281,167 training images, 50,000 validation images, and 100,000 test images across 1,000 object classes organized by the WordNet hierarchy. Each image is human-annotated and quality-controlled. It has been the standard evaluation dataset for deep learning vision models since the AlexNet breakthrough in 2012 and remains a foundational benchmark in computer vision.
Keywords
vision image-classification benchmark imagenet wordnet deep-learning ilsvrc
Topics
Vision / Image Classification
Research notes
- Requires agreeing to ImageNet Terms of Access on HuggingFace. Non-commercial research and educational use only.