Logo do repositório

Hate Speech Detection in Portuguese Using BERTimbau

Carregando...
Imagem de Miniatura

Orientador

Coorientador

Pós-graduação

Curso de graduação

Título da Revista

ISSN da Revista

Título de Volume

Editor

Tipo

Trabalho apresentado em evento

Direito de acesso

Resumo

Hate speech refers to language expressions that attack individuals or groups based on specific characteristics associated with their identities, causing lasting damage. Social networks have become a pertinent environment for hate speech proliferation since they allow anonymity and maintain a safe distance from aggressors and assaulted victims. With the amount of data published every minute, automatic identification of hate speech using machine learning gathered much attention from academic and industrial researchers. However, as with many natural language processing tasks, the efforts mainly focused on English, and languages like Portuguese remain less explored. Therefore, this paper aims to experiment with different techniques to deal with the challenges associated with low-resource languages in automatic hate speech detection. It evaluates whether knowledge transferred from offensive speech detection as a source task can be effective for hate detection and if the unbalanced data poses an obstacle for a Portuguese pre-trained BERT model, BERTimbau. Experimental results show that transferring learning between tasks does not improve performance and that using balanced data leads to better F1 scores and Cohen’s Kappa.

Descrição

Palavras-chave

Hate Speech, Machine Learning, Natural Language Processing, Portuguese Language, Undersampling

Idioma

Inglês

Citação

Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics), v. 15368 LNCS, p. 244-255.

Itens relacionados

Financiadores

Unidades

Item type:Unidade,
Faculdade de Ciências
FC
Campus: Bauru


Departamentos

Cursos de graduação

Programas de pós-graduação

Outras formas de acesso