DravidianCodeMix: sentiment analysis and offensive language identification dataset for Dravidian languages in code-mixed text.
This paper describes the development of a multilingual, manually annotated dataset for three under-resourced Dravidian languages generated from social media comments. The dataset was annotated for sentiment analysis and offensive language identification for a total of more than 60,000 YouTube commen...
| Published in: | Language Resources & Evaluation Vol. 56; no. 3; pp. 765 - 807 |
|---|---|
| Main Authors: | , , , , , , |
| Format: | Article |
| Published: |
Springer Nature
Sep2022
|
| Subjects: | |
| Online Access: | View this record in EBSCOhost |