MultiDomainYT-ToxBR: Um Corpus Multidomínio de Comentários Tóxicos do YouTube em Português Brasileiro

  • Gabriel Z. de Souza PUCRS
  • Isabel H. Manssour PUCRS

Resumo


Este trabalho apresenta o MultiDomainYT-ToxBR, um corpus de 5.000 comentários do YouTube anotados em nove categorias de toxicidade e distribuídos em cinco domínios temáticos no contexto brasileiro. A análise de concordância entre anotadores evidencia desafios inerentes à tarefa, com índices moderados em categorias como racismo e xenofobia, cujas manifestações veladas e dependentes de contexto dificultam o consenso. A comparação com o ToLD-Br revela padrões complementares: o novo corpus supera o anterior em concordância para homofobia (α = 0,73 vs. 0,68) e introduz categorias inéditas como intolerância política e religiosa. Como validação, o BERTweet.BR, após fine-tuning, atingiu F1 de 0,78 na classificação binária.

Referências

Assis, G., Amorim, A., Carvalho, J., de Oliveira, D., Vianna, D., and Paes, A. (2024). Exploring Portuguese hate speech detection in low-resource settings: Lightly tuning encoder models or in-context learning of large models? In Gamallo, P., Claro, D., Teixeira, A., Real, L., Garcia, M., Oliveira, H. G., and Amaro, R., editors, Proceedings of the 16th International Conference on Computational Processing of Portuguese - Vol. 1, pages 301–311, Santiago de Compostela, Galicia/Spain. Association for Computational Lingustics.

Carneiro, F., Vianna, D., Carvalho, J., Plastino, A., and Paes, A. (2025). BERTweet.BR: a pre-trained language model for tweets in Portuguese. Neural Computing and Applications, 37(6):4363–4385.

de Pelle, R. P. and Moreira, V. P. (2017). Offensive comments in the brazilian web: a dataset and baseline results.

Fortuna, P., Rocha da Silva, J., Soler-Company, J., Wanner, L., and Nunes, S. (2019). A hierarchically-labeled Portuguese hate speech dataset. In Roberts, S. T., Tetreault, J., Prabhakaran, V., and Waseem, Z., editors, Proceedings of the Third Workshop on Abusive Language Online, pages 94–104, Florence, Italy. Association for Computational Linguistics.

Junqueira, J. d. R., Lopes, , M., C. L. D. S., da Silva, F. L. V., Carvalho, E. A., de Freitas, L. A., and Corrêa, U. B. (2024). Sabiá in action: An investigation of its abilities in aspect-based sentiment analysis, hate speech detection, irony detection, and question-answering. In IJCNN, pages 1–8.

Landis, J. R. and Koch, G. G. (1977). The measurement of observer agreement for categorical data. Biometrics, 33(1):159–174.

Leite, J. A., Silva, D., Bontcheva, K., and Scarton, C. (2020). Toxic language detection in social media for Brazilian Portuguese: New dataset and multilingual analysis. In Wong, K.-F., Knight, K., and Wu, H., editors, Proceedings of the 1st Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics and the 10th International Joint Conference on Natural Language Processing, pages 914–924, Suzhou, China. Association for Computational Linguistics.

Luo, C., Bhambhoria, R., Dahan, S., and Zhu, X. (2023). Legally enforceable hate speech detection for public forums. In Bouamor, H., Pino, J., and Bali, K., editors, Findings of the Association for Computational Linguistics: EMNLP 2023, pages 10948–10963, Singapore. Association for Computational Linguistics.

Milli, S., Carroll, M., Wang, Y., Pandey, S., Zhao, S., and Dragan, A. D. (2025). Engagement, user satisfaction, and the amplification of divisive content on social media. PNAS Nexus, 4(3):pgaf062.

Müller, K. and Schwarz, C. (2021). Fanning the flames of hate: Social media and hate crime. Journal of the European Economic Association, 19(4):2131–2167.

Oliveira, A., Silva, P. H., Santos, V., Moreira, G., Freitas, V. L., and Luz, E. J. (2024a). Toxic text classification in portuguese: Is llama 3.1 8b all you need? In Anais do XV Simpósio Brasileiro de Tecnologia da Informação e da Linguagem Humana, pages 57–66, Porto Alegre, RS, Brasil. SBC.

Oliveira, A. d. S., Cecote, T. d. C., Alvarenga, J. P. R., de Souza Freitas, V. L., and da Silva Luz, E. J. (2024b). Toxic speech detection in Portuguese: A comparative study of large language models. In Gamallo, P., Claro, D., Teixeira, A., Real, L., Garcia, M., Oliveira, H. G., and Amaro, R., editors, Proceedings of the 16th Internati onal Conference on Computational Processing of Portuguese - Vol. 1, pages 108–116, Santiago de Compostela, Galicia/Spain. Association for Computational Lingustics.

Ortellado, P., Ribeiro, M. M., and Zeine, L. (2022). Existe polarização política no brasil? análise das evidências em duas séries de pesquisas de opinião. Opinião Pública, 28(1):62–91.

Trajano, D., Bordini, R. H., and Vieira, R. (2024). Olid-br: offensive language identification dataset for brazilian portuguese. Language Resources and Evaluation, 58(4):1263–1289.

Trajano, D. d. O. (2023). Detecção de linguagem tóxica aplicada a textos em português. Dissertação (mestrado em ciência da computação), Pontifícia Universidade Católica do Rio Grande do Sul, Porto Alegre, RS, Brasil. Orientador: Rafael Heitor Bordini.

Vargas, F., Carvalho, I., Rodrigues de Góes, F., Pardo, T., and Benevenuto, F. (2022). HateBR: A large expert annotated corpus of Brazilian Instagram comments for offensive language and hate speech detection. In Calzolari, N., Béchet, F., Blache, P., Choukri, K., Cieri, C., Declerck, T., Goggi, S., Isahara, H., Maegaard, B., Mariani, J., Mazo, H., Odijk, J., and Piperidis, S., editors, Proceedings of the Thirteenth Language Resources and Evaluation Conference, pages 7174–7183, Marseille, France. European Language Resources Association.

Zufall, F., Hamacher, M., Kloppenborg, K., and Zesch, T. (2022). A legal approach to hate speech – operationalizing the EU’s legal framework against the expression of hatred as an NLP task. In Aletras, N., Chalkidis, I., Barrett, L., Goant, ă, C., and Preot,iuc-Pietro, D., editors, Proceedings of the Natural Legal Language Processing Workshop 2022, pages 53–64, Abu Dhabi, United Arab Emirates (Hybrid). Association for Computational Linguistics.
Publicado
19/10/2026
SOUZA, Gabriel Z. de; MANSSOUR, Isabel H.. MultiDomainYT-ToxBR: Um Corpus Multidomínio de Comentários Tóxicos do YouTube em Português Brasileiro. In: SIMPÓSIO BRASILEIRO DE TECNOLOGIA DA INFORMAÇÃO E DA LINGUAGEM HUMANA (STIL), 17. , 2026, Cuiabá/MT. Anais [...]. Porto Alegre: Sociedade Brasileira de Computação, 2026 . p. 441-453. DOI: https://doi.org/10.5753/stil.2026.26542.