Intelligence

ai

Detoxifying Toxic Communication: A Design Science Approach to Responsible AI

arXiv: Computers and SocietyInternationalHigh confidence1 min

What changed

Research from arXiv explores a Design Science approach to developing a responsible AI artifact aimed at addressing toxic communication in digital workplaces. Current moderation tools often disrupt communication by blocking or deleting content. This new approach integrates transformer-based classifiers for detecting toxic language with a generative detoxification model that rewrites such text into semantically equivalent, non-offensive paraphrases, demonstrating high accuracy in detection and strong semantic preservation.

Why it matters

The development of AI solutions capable of not just detecting but also 'detoxifying' harmful communication represents a significant advancement in fostering healthier digital environments. This shifts from purely punitive moderation to a more constructive approach, which can preserve dialogue while mitigating its negative impacts.

What to watch

Toxic language in digital workplaces, encompassing pejoratives, sarcasm, condescension, and subtle incivility, negatively impacts trust, morale, and collaboration.

Forward consideration, not a verified fact.

Reported by arXiv: Computers and Society, International. The document itself is not reproduced here.

Read the original publication