Aller au contenu principal
Accès ouvert déclaré2025article

The unappreciated role of intent in algorithmic moderation of abusive content on social media

7Citations signalées
2Institutions associées
1Pays d’affiliation

Résumé fourni par la source

A significant body of research is dedicated to developing language models that can detect various types of online abuse, for example, hate speech, cyberbullying. However, there is a disconnect between platform policies, which often consider the author's intention as a criterion for content moderation, and the current capabilities of detection models, which typically lack efforts to capture intent. This paper examines the role of intent in the moderation of abusive content. Specifically, we review state-of-the-art detection models and benchmark training datasets to assess their ability to capture intent. We propose changes to the design and development of automated detection and moderation systems to improve alignment with ethical and policy conceptualizations of these abuses.

Institutions

Sujets associés

Hate Speech and Cyberbullying DetectionBullying, Victimization, and AggressionCybercrime and Law Enforcement Studies

BNTIC News n’est pas le producteur de ces données. Métadonnées interrogées à la demande auprès de OpenAlex (CC0). Sources et limites.