DravidianCodeMix: sentiment analysis and offensive language identification dataset for Dravidian languages in code-mixed text.
This paper describes the development of a multilingual, manually annotated dataset for three under-resourced Dravidian languages generated from social media comments. The dataset was annotated for sentiment analysis and offensive language identification for a total of more than 60,000 YouTube commen...
| Publicado en: | Language Resources & Evaluation Vol. 56; no. 3; pp. 765 - 807 |
|---|---|
| Autores principales: | , , , , , , |
| Formato: | Artículo |
| Publicado: |
Springer Nature
Sep2022
|
| Materias: | |
| Acceso en línea: | Ver este registro en EBSCOhost |