30 repos
NLP systems for detecting abusive and offensive content across multiple languages and code-mixed text, primarily using transformer-based models like BERT and T5. The cluster focuses on text classification approaches for identifying harmful language in English, Hindi, Bengali, Urdu, Kannada, and related language variants, with applications to content moderation and safer online spaces. Most repos center on dataset creation, model training, and evaluation frameworks for these multilingual and code-switched abuse detection tasks.