**arXiv ID:** 2109.03552 **Authors:** Saurabh Gaikwad, Tharindu Ranasinghe, Marcos Zampieri, Christopher M. Homan **Published:** 2021-09-08T11:29:44Z **Abstract:** The widespread presence of offensive language on social media motivated the development of systems capable of recognizing such content automatically. Apart from a few notable exceptions, most research on automatic offensive language identification has dealt with English. To address this shortcoming, we introduce MOLD, the Marathi O...
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill crosslingual-offensive-language-identification-for-low-resource-languages-the-case-of-marathi --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Crosslingual Offensive Language Identification For Low Resource Languages The Case Of Marathi?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-crosslingual-offensive-language-identification-for)More formats (shields.io, HTML) on the badges page.
# Cross-lingual Offensive Language Identification for Low Resource Languages: The Case of Marathi
**arXiv ID:** 2109.03552
**Authors:** Saurabh Gaikwad, Tharindu Ranasinghe, Marcos Zampieri, Christopher M. Homan
**Published:** 2021-09-08T11:29:44Z
**Abstract:**
The widespread presence of offensive language on social media motivated the development of systems capable of recognizing such content automatically. Apart from a few notable exceptions, most research on automatic offensive language identification has dealt with English. To address this shortcoming, we introduce MOLD, the Marathi Offensive Language Dataset. MOLD is the first dataset of its kind compiled for Marathi, thus opening a new domain for research in low-resource Indo-Aryan languages. We present results from several machine learning experiments on this dataset, including zero-short and other transfer learning experiments on state-of-the-art cross-lingual transformers from existing data in Bengali, English, and Hindi.
## Skill Description
This skill is generated from the arXiv paper: Cross-lingual Offensive Language Identification for Low Resource Languages: The Case of Marathi (2109.03552).
## How to Use
[To be filled in by the user or by future automation]
## References
- [arXiv:2109.03552](http://arxiv.org/abs/2109.03552v1)
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!