Overview of JADE 3.0 Safety Alignment Dataset
mainJADE 3.0 is a high-quality Chinese dataset designed for the safety fine-tuning of Large Language Models (LLMs). It consists of triplets containing a "High-risk Question", a "Violating Response", and a "Safe and Useful Response". These triplets are generated based on high-risk scenarios identified by the JADE 1.0 testing platform, using LLM self-reflection and correction to ensure high quality.
Note: The dataset contains examples of harmful and violating content for training purposes and does not represent the stance of the development team.