Knowledge Resource · Open access
Detoxifying Toxic Communication: A Design Science Approach to Responsible AI
- Author
- Aziz Shuaib Ausi
- Published
- 8 September 2026
- Reading time
- 1 min
- Publication type
- Knowledge Resource
- Availability
- Open access
Research from arXiv explores a Design Science approach to developing a responsible AI artifact aimed at addressing toxic communication in digital workplaces. Current moderation tools often disrupt communication by blocking or deleting content. This new approach integrates transformer-based classifiers for detecting toxic language with a generative detoxification model that rewrites such text into semantically equivalent, non-offensive paraphrases, demonstrating high accuracy in detection and strong semantic preservation.
Why it matters
The development of AI solutions capable of not just detecting but also 'detoxifying' harmful communication represents a significant advancement in fostering healthier digital environments. This shifts from purely punitive moderation to a more constructive approach, which can preserve dialogue while mitigating its negative impacts.
Key insights
- Toxic language in digital workplaces, encompassing pejoratives, sarcasm, condescension, and subtle incivility, negatively impacts trust, morale, and collaboration.
- Existing communication moderation tools typically focus on deleting or blocking harmful messages, which can disrupt communication flow and do not offer constructive resolution.
- A Design Science Research approach was used to create a responsible AI artifact for detecting and detoxifying toxic communication.
- The artifact combines fine-tuned transformer-based classifiers (DistilBERT, DistilRoBERTa) for toxicity detection.
- It utilizes a generative detoxification model (mT0-XL-Detox-ORPO) to rewrite toxic text into non-offensive, semantically equivalent paraphrases.
- Technical evaluations indicate high accuracy in toxicity detection and robust semantic preservation during the detoxification process.
Source
arXiv — Computers and Society — https://arxiv.org/abs/2609.00361
Related resources
Previous
A Mathematical Framework for Legacy, Governance, and Decision Integrity in Enterprise AI
Next
Qualified Cross-References as a Verification Method: The Normative Environment of the EU AI Act
Social bots weaken activist cohesion
Knowledge Resource
How Does LGBTQIA+ Identity Affect LLM Behavior? Implications for Requirements Engineering of Mental Health AI Systems
Knowledge Resource
Qualified Cross-References as a Verification Method: The Normative Environment of the EU AI Act
Knowledge Resource
A Mathematical Framework for Legacy, Governance, and Decision Integrity in Enterprise AI
Knowledge Resource
Towards AI-Assisted Clinical Trial Matching: Practical Considerations, Multicenter Evaluation, and Real-World Deployment
Knowledge Resource
The Five Safes as a Privacy Context
Knowledge Resource
Citation
Cite this publication (APA 7)
Aziz Shuaib Ausi (2026). Detoxifying Toxic Communication: A Design Science Approach to Responsible AI. Knowledge Resource. Aziz Shuaib Ausi. https://www.azizshuaib.com/verify/ASA-EXE-2026-00265
Verification
This is an authenticated institutional record.
- Verification ID
- ASA-EXE-2026-00265
- Version
- v1.0 · r0
- Issued
- 8 September 2026
- Publisher
- Aziz Shuaib Ausi
- Licence
- All rights reserved. Reproduction requires written permission.