XDetox: Text Detoxification with Token-Level Toxicity Explanations
- Authors
- Lee, Beomseok; Kim, Hyunwoo; Kim, Keon; Choi, Yong Suk
- Issue Date
- Nov-2024
- Publisher
- Association for Computational Linguistics (ACL)
- Citation
- EMNLP 2024 - 2024 Conference on Empirical Methods in Natural Language Processing, Proceedings of the Conference, pp 15215 - 15226
- Pages
- 12
- Indexed
- SCOPUS
- Journal Title
- EMNLP 2024 - 2024 Conference on Empirical Methods in Natural Language Processing, Proceedings of the Conference
- Start Page
- 15215
- End Page
- 15226
- URI
- https://scholarworks.bwise.kr/hanyang/handle/2021.sw.hanyang/206715
- DOI
- 10.18653/v1/2024.emnlp-main.848
- Abstract
- Methods for mitigating toxic content through masking and infilling often overlook the decision-making process, leading to either insufficient or excessive modifications of toxic tokens. To address this challenge, we propose XDetox, a novel method that integrates token-level toxicity explanations with the masking and infilling detoxification process. We utilized this approach with two strategies to enhance the performance of detoxification. First, identifying toxic tokens to improve the quality of masking. Second, selecting the regenerated sentence by re-ranking the least toxic sentence among candidates. Our experimental results show state-of-the-art performance across four datasets compared to existing detoxification methods. Furthermore, human evaluations indicate that our method outperforms baselines in both fluency and toxicity reduction. These results demonstrate the effectiveness of our method in text detoxification.
- Files in This Item
-
Go to Link
- Appears in
Collections - 서울 공과대학 > 서울 컴퓨터소프트웨어학부 > 1. Journal Articles

Items in ScholarWorks are protected by copyright, with all rights reserved, unless otherwise indicated.