DravidianCodeMix: sentiment analysis and offensive language identification dataset for Dravidian languages in code-mixed text.

This paper describes the development of a multilingual, manually annotated dataset for three under-resourced Dravidian languages generated from social media comments. The dataset was annotated for sentiment analysis and offensive language identification for a total of more than 60,000 YouTube commen...

Descripción completa

Detalles Bibliográficos
Publicado en:Language Resources & Evaluation Vol. 56; no. 3; pp. 765 - 807
Autores principales: Chakravarthi, Bharathi Raja, Priyadharshini, Ruba, Muralidaran, Vigneshwaran, Jose, Navya, Suryawanshi, Shardul, Sherly, Elizabeth, McCrae, John P.
Formato: Artículo
Publicado: Springer Nature Sep2022
Materias:
Acceso en línea:Ver este registro en EBSCOhost