[EMNLP 2024] PREDICT: Multi-Agent-based Debate Simulation for Generalized Hate Speech Detection

Hanyang Human-Centered Computing Lab · Advanced ·🤖 AI Agents & Automation ·1y ago

About this lesson

PREDICT: Multi-Agent-based Debate Simulation for Generalized Hate Speech Detection (EMNLP '24) Someen Park, Jaehoon Kim, Seungwan Jin, Sohyun Park, Kyungsik Han Paper: https://aclanthology.org/2024.emnlp-main.1166/ Abstract: While a few public benchmarks have been proposed for training hate speech detection models, the differences in labeling criteria between these benchmarks pose challenges for generalized learning, limiting the applicability of the models. Previous research has presented methods to generalize models through data integration or augmentation, but overcoming the differences in labeling criteria between datasets remains a limitation. To address these challenges, we propose PREDICT, a novel framework that uses the notion of multi-agent for hate speech detection. PREDICT consists of two phases: (1) PRE (Perspective-based REasoning): Multiple agents are created based on the induced labeling criteria of given datasets, and each agent generates stances and reasons; (2) DICT (Debate using InCongruenT references): Agents representing hate and non-hate stances conduct the debate, and a judge agent classifies hate or non-hate and provides a balanced reason. Experiments on five representative public benchmarks show that PREDICT achieves superior cross-evaluation performance compared to methods that focus on specific labeling criteria or majority voting methods. Furthermore, we validate that PREDICT effectively mediates differences between agents’ opinions and appropriately incorporates minority opinions to reach a consensus. Our code is available at https://github.com/Hanyang-HCC-Lab/PREDICT

Original Description

PREDICT: Multi-Agent-based Debate Simulation for Generalized Hate Speech Detection (EMNLP '24) Someen Park, Jaehoon Kim, Seungwan Jin, Sohyun Park, Kyungsik Han Paper: https://aclanthology.org/2024.emnlp-main.1166/ Abstract: While a few public benchmarks have been proposed for training hate speech detection models, the differences in labeling criteria between these benchmarks pose challenges for generalized learning, limiting the applicability of the models. Previous research has presented methods to generalize models through data integration or augmentation, but overcoming the differences in labeling criteria between datasets remains a limitation. To address these challenges, we propose PREDICT, a novel framework that uses the notion of multi-agent for hate speech detection. PREDICT consists of two phases: (1) PRE (Perspective-based REasoning): Multiple agents are created based on the induced labeling criteria of given datasets, and each agent generates stances and reasons; (2) DICT (Debate using InCongruenT references): Agents representing hate and non-hate stances conduct the debate, and a judge agent classifies hate or non-hate and provides a balanced reason. Experiments on five representative public benchmarks show that PREDICT achieves superior cross-evaluation performance compared to methods that focus on specific labeling criteria or majority voting methods. Furthermore, we validate that PREDICT effectively mediates differences between agents’ opinions and appropriately incorporates minority opinions to reach a consensus. Our code is available at https://github.com/Hanyang-HCC-Lab/PREDICT
Watch on YouTube ↗ (saves to browser)
Sign in to unlock AI tutor explanation · ⚡30

Related Reads

📰
How I Turned a Raspberry Pi Into an AI-Powered Automation Hub
Learn how to turn a Raspberry Pi into an AI-powered automation hub, enabling automated tasks and workflows
Dev.to · ULNIT
📰
Building an Incident Triage Agent with Full Observability in SigNoz
Learn to build an incident triage agent with full observability in SigNoz to overcome AI black boxes
Dev.to · Aadvik Krishna
📰
I Built a Nutrition-Scoring Bot Inside Telegram, and Swiggy Just Approved It
Learn how to build a nutrition-scoring bot inside Telegram and deploy it to production, as demonstrated by a successful integration with Swiggy
Dev.to · Magithar Sridhar
📰
The Greatest Transformation of Future Civilization Is Not AI, but Organization
The future of civilization will be shaped by advancements in organization, not just AI, and leaders must adapt to coordinate intelligence and protocols, not just manage people.
Medium · AI
Up next
This Rust AI Agent Is Insanely Fast #coding #ai
Silism
Watch →