Detoxifying Toxic Communication: A Design Science Approach to Responsible AI
cs.CY, cs.CL
Submitted: 2026-08-31
Updated: 2026-08-31
Code: https://github.com/thearhamn/detox_org_com
Terminology
Sources
- Style Transfer as Data Augmentation: A Case Study on Named Entity Recognition
- Optimus: A Robust Defense Framework for Mitigating Toxicity while Fine-Tuning Conversational AI
- Text Detoxification using Large Pre-trained Neural Models
- ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
- The Disagreement Problem in Explainable Machine Learning: A Practitioner's Perspective
- RoBERTa: A Robustly Optimized BERT Pretraining Approach
- Toxicity Detection: Does Context Really Matter?
- Directions in Abusive Language Training Data: Garbage In, Garbage Out
Related papers
- Reasoning Enhances Robustness to Prompt Injection in LLM-Based Consensus
- Generative AI Purpose-built for Social and Mental Health: A Real-World Pilot
- PersonaMem-v3: Toward Omni-Platform Personal Intelligence for Holistic User Understanding, Recommendation, and Agentic Tasks
- What is an intelligent system?
- AI University: An LLM-Powered Learning Assistant for Engineering---A Finite Element Method Case Study
- Generative AI Use in Entrepreneurship: An Integrative Review and an Empowerment-Entrapment Framework