Léo Laugier scite author profile

Léo Laugier

5Publications

60Citation Statements Received

118Citation Statements Given

How they've been cited

How they cite others

117

Affiliations

Publications

Order By: Most citations

Civil Rephrases Of Toxic Texts With Self-Supervised Transformers

Laugier¹,

Pavlopoulos²,

Sorensen³

et al. 2021

View full text Add to dashboard Cite

Platforms that support online commentary, from social networks to news sites, are increasingly leveraging machine learning to assist their moderation efforts. But this process does not typically provide feedback to the author that would help them contribute according to the community guidelines. This is prohibitively time-consuming for human moderators to do, and computational approaches are still nascent. This work focuses on models that can help suggest rephrasings of toxic comments in a more civil manner. Inspired by recent progress in unpaired sequence-tosequence tasks, a self-supervised learning model is introduced, called CAE-T5 1 . CAE-T5 employs a pre-trained text-to-text transformer, which is fine tuned with a denoising and cyclic auto-encoder loss. Experimenting with the largest toxicity detection dataset to date (Civil Comments) our model generates sentences that are more fluent and better at preserving the initial content compared to earlier text style transfer systems which we compare with using several scoring systems and human evaluation.

show abstract

SemEval-2021 Task 5: Toxic Spans Detection

Pavlopoulos¹,

Sorensen²,

Laugier³

et al. 2021

View full text Add to dashboard Cite

The Toxic Spans Detection task of SemEval-2021 required participants to predict the spans of toxic posts that were responsible for the toxic label of the posts. The task could be addressed as supervised sequence labeling, using training data with gold toxic spans provided by the organisers. It could also be treated as rationale extraction, using classifiers trained on potentially larger external datasets of posts manually annotated as toxic or not, without toxic span annotations. For the supervised sequence labeling approach and evaluation purposes, posts previously labeled as toxic were crowd-annotated for toxic spans. Participants submitted their predicted spans for a held-out test set, and were scored using character-based F1. This overview summarises the work of the 36 teams that provided system descriptions.

show abstract

From the Detection of Toxic Spans in Online Discussions to the Analysis of Toxic-to-Civil Transfer

Pavlopoulos¹,

Laugier²,

Xenos³

et al. 2022

View full text Add to dashboard Cite

We study the task of toxic spans detection, which concerns the detection of the spans that make a text toxic, when detecting such spans is possible. We introduce a dataset for this task, TOXICSPANS, which we release publicly. By experimenting with several methods, we show that sequence labeling models perform best. Moreover, methods that add generic rationale extraction mechanisms on top of classifiers trained to predict if a post is toxic or not are also surprisingly promising. Finally, we use TOXICSPANS and systems trained on it, to provide further analysis of state-of-the-art toxic to non-toxic transfer systems, as well as of human performance on that latter task. Our work highlights challenges in finer toxicity detection and mitigation.

show abstract

Civil Rephrases Of Toxic Texts With Self-Supervised Transformers

Laugier¹,

Pavlopoulos²,

Sorensen³

et al. 2021

Preprint

View full text Add to dashboard Cite

Harmful Language Datasets: An Assessment of Robustness

Korre¹,

Pavlopoulos²,

Sorensen³

et al. 2023

View full text Add to dashboard Cite

The automated detection of harmful language has been of great importance for the online world, especially with the growing importance of social media and, consequently, polarisation.There are many open challenges to high quality detection of harmful text, from dataset creation to generalisable application, thus calling for more systematic studies. In this paper, we explore re-annotation as a means of examining the robustness of already existing labelled datasets, showing that, despite using alternative definitions, the inter-annotator agreement remains very inconsistent, highlighting the intrinsically subjective and variable nature of the task. In addition, we build automatic toxicity detectors using the existing datasets, with their original labels, and we evaluate them on our multi-definition and multi-source datasets. Surprisingly, while other studies show that hate speech detection models perform better on data that are derived from the same distribution as the training set, our analysis demonstrates this is not necessarily true.

show abstract

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

hi@scite.ai

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

Léo Laugier

Civil Rephrases Of Toxic Texts With Self-Supervised Transformers

SemEval-2021 Task 5: Toxic Spans Detection

From the Detection of Toxic Spans in Online Discussions to the Analysis of Toxic-to-Civil Transfer

Civil Rephrases Of Toxic Texts With Self-Supervised Transformers

Harmful Language Datasets: An Assessment of Robustness

Contact Info

Product

Resources

About