← Back to explorer

Uncensored-SFT-v2 (High Quality Uncensored Instruction Dataset V2)

Type
dataset
Venue
Hugging Face
Year
2026
Source
huggingface
Access
free
Language
English
Added
2026-07-17T20:18:03.812477+00:00
Verified
2026-07-17T20:18:03.812477+00:00

Summary

A semantically deduplicated instruction-tuning dataset of ~485k input/output pairs covering uncensored, harmful, and adversarial prompts (scams, hacking, explicit content, etc.) with compliant responses. It is intended for training 'uncensored' LLMs that do not refuse requests and for safety/alignment research on model behavior. V2 is a deduplicated version of the original V1 release.

Keywords

uncensored sft instruction-tuning harmful-prompts adversarial red-teaming safety-research alignment

Topics

NLP / Instruction Tuning

Research notes

  • Contains explicit, harmful, and illegal-activity content. Intended for research purposes only; not for deployment without safety considerations.