Detection and Safeguarding of Chinese Toxic Content from Users and LLMs: A Survey
This paper presents a comprehensive survey that systematically reviews research on detecting user-generated toxic content on Chinese social media and safeguarding Chinese large language models against LLM-generated toxicity, aiming to consolidate fragmented progress and identify key challenges for future development.