CL LGJan 8, 2024

Overview of the 2023 ICON Shared Task on Gendered Abuse Detection in Indic Languages

Aatman Vaidya, Arnav Arora, Aditya Joshi, Tarunima Prabhakar

arXiv:2401.03677v11.0h-index: 12

Originality Synthesis-oriented

AI Analysis

It addresses the problem of online gendered abuse detection for Indic language communities, but is incremental as it builds on existing shared task frameworks.

The paper reports on the 2023 ICON shared task for detecting gendered abuse in Hindi, Tamil, and Indian English, using a dataset of about 6,500 training and 1,200 test posts, with best F-1 scores ranging from 0.572 to 0.616 across subtasks.

This paper reports the findings of the ICON 2023 on Gendered Abuse Detection in Indic Languages. The shared task deals with the detection of gendered abuse in online text. The shared task was conducted as a part of ICON 2023, based on a novel dataset in Hindi, Tamil and the Indian dialect of English. The participants were given three subtasks with the train dataset consisting of approximately 6500 posts sourced from Twitter. For the test set, approximately 1200 posts were provided. The shared task received a total of 9 registrations. The best F-1 scores are 0.616 for subtask 1, 0.572 for subtask 2 and, 0.616 and 0.582 for subtask 3. The paper contains examples of hateful content owing to its topic.

View on arXiv PDF

Similar