Identificação de Expressões Anafóricas e Não Anafóricas com Base na Estrutura do Sintagma
Resumo
Um dos problemas da resolução de anáforas é identificar quais expressões são anafóricas e quais são não anafóricas. Neste trabalho um conjunto de heurísticas para identificar as expressões não anafóricas, a implementação destas heurísticas em um ambiente para a resolução de anáforas (ART - Anaphor Resolution Tool) e uma avaliação dos resultados obtidos são apresentados.
Referências
Bean, D. L. and Riloff, E. (1999) “Corpus-based Identification of Non-Anaphoric Noun Phrases”. In: Proceedings of the 37th Annual Meeting of the Association for Computational Linguistics, p. 373–380.
Bick, E. (2000) “The Parsing System PALAVRAS: Automatic Grammatical Analysis of Portuguese in a Constraint Grammar Framework”. PhD thesis, Arhus University, Arhus.
Cardie, C. and Wagstaff, K. (1999) “Noun phrase coreference as clustering”. In: Proceedings of the 1999 SIGDAT Conference on Empirical Methods in Natural Language Processing and Very Large Corpora, College Park, p. 82–89.
Gasperin, C.; Vieira, R.; Goulart, R.; Quaresma, P. (2003) “Extrating XML Syntactic Chunks from Portuguese Corpora”. In: Traitement Automatique Dês Langues Minoritaires-TALN, Btaz-sur-mer, France.
Gasperin, C., Goulart, R.; Vieira, R. (2003) “Uma Ferramenta para Resolução Automática de Co-referência”. Anais do Encontro Nacional de Inteligência Artificial (ENIA 2003), Campinas, SP.
McCarthy, J. F. and Lehnert, G. (1995) “Using decision trees for coreference resolution”. In: Proceedings of the 14th International Joint Conference on Artificial Intelligence, Montreal, Canada, p. 1050–1055.
Miorelli, S. (2001) “Extração do Sintagma Nominal em Sentenças em Português”. Dissertação de Mestrado, PUC, Porto Alegre.
Müller, C. and Strube, M. (2000) “MMAX: A tool for the annotation of multi-modal corpora”. In: Proceedings of the IJCAI 2001, Seattle, p. 45–50.
Muller, C.; Stefan, R.; Strube, M. (2002) “Applying Co-training to reference resolution”. In: Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics (ACL-2002), Philadelphia, Penn., p. 352-359.
Ng, V. and Cardie, C. (2002a) “Improving machine learning approaches to coreference resolution”. In: Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics.
Ng, V. and Cardie, C. (2002b) “Identifying Anaphoric and Non-Anaphoric Noun Phrases to Improve Coreference Resolution”. In: Proceedings of the Nineteenth International Conference on Computational Linguistics (COLING-2002), p. 730–736.
Perini, M. (2003) Gramática descritiva do português. São Paulo: Editora Ática, 380 p.
Silva, M. and Koch, I. (1989). Lingüística Aplicada ao Português: Sintaxe. São Paulo: Editora Cortez, 160 p.
Soon, W. M.; Ng, H.wee T.; Lim, D. C. Y. (2001) “A machine learning approach to coreference resolution of noun phrases”. In: Computational Linguistics, p. 521–544.
Uryupina, O. (2003) “High-precision Identification of Discourse New and Unique Noun Phrases”. In: Proceedings of the ACL Student Workshop, Sapporo.
Vieira, R. (1998) “Definite description processing in unrestricted text”. PhD thesis, University of Edinburgh, Edinburgh.
Vieira, V. and Poesio, M. (2000) “An empirically-based system for processing definite descriptions”. In: Computational Linguistics.
Vieira, R.; Salmon-Alt, S.; Schang, E. (2002) “Multilingual corpora annotation for processing definite descriptions”. In: Proceedings of the PorTAL 2002, Faro.
Vieira, R.; Gasperin, C.; Goulart, R.; Salmon-Alt, S. (2003) “From concrete to virtual annotation mark-up language: the case of COMMOn-REFs”. In Proceedings of the (ACL 2003) Workshop on Linguistic Annotation: Getting the Model Right, Sapporo.
Bick, E. (2000) “The Parsing System PALAVRAS: Automatic Grammatical Analysis of Portuguese in a Constraint Grammar Framework”. PhD thesis, Arhus University, Arhus.
Cardie, C. and Wagstaff, K. (1999) “Noun phrase coreference as clustering”. In: Proceedings of the 1999 SIGDAT Conference on Empirical Methods in Natural Language Processing and Very Large Corpora, College Park, p. 82–89.
Gasperin, C.; Vieira, R.; Goulart, R.; Quaresma, P. (2003) “Extrating XML Syntactic Chunks from Portuguese Corpora”. In: Traitement Automatique Dês Langues Minoritaires-TALN, Btaz-sur-mer, France.
Gasperin, C., Goulart, R.; Vieira, R. (2003) “Uma Ferramenta para Resolução Automática de Co-referência”. Anais do Encontro Nacional de Inteligência Artificial (ENIA 2003), Campinas, SP.
McCarthy, J. F. and Lehnert, G. (1995) “Using decision trees for coreference resolution”. In: Proceedings of the 14th International Joint Conference on Artificial Intelligence, Montreal, Canada, p. 1050–1055.
Miorelli, S. (2001) “Extração do Sintagma Nominal em Sentenças em Português”. Dissertação de Mestrado, PUC, Porto Alegre.
Müller, C. and Strube, M. (2000) “MMAX: A tool for the annotation of multi-modal corpora”. In: Proceedings of the IJCAI 2001, Seattle, p. 45–50.
Muller, C.; Stefan, R.; Strube, M. (2002) “Applying Co-training to reference resolution”. In: Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics (ACL-2002), Philadelphia, Penn., p. 352-359.
Ng, V. and Cardie, C. (2002a) “Improving machine learning approaches to coreference resolution”. In: Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics.
Ng, V. and Cardie, C. (2002b) “Identifying Anaphoric and Non-Anaphoric Noun Phrases to Improve Coreference Resolution”. In: Proceedings of the Nineteenth International Conference on Computational Linguistics (COLING-2002), p. 730–736.
Perini, M. (2003) Gramática descritiva do português. São Paulo: Editora Ática, 380 p.
Silva, M. and Koch, I. (1989). Lingüística Aplicada ao Português: Sintaxe. São Paulo: Editora Cortez, 160 p.
Soon, W. M.; Ng, H.wee T.; Lim, D. C. Y. (2001) “A machine learning approach to coreference resolution of noun phrases”. In: Computational Linguistics, p. 521–544.
Uryupina, O. (2003) “High-precision Identification of Discourse New and Unique Noun Phrases”. In: Proceedings of the ACL Student Workshop, Sapporo.
Vieira, R. (1998) “Definite description processing in unrestricted text”. PhD thesis, University of Edinburgh, Edinburgh.
Vieira, V. and Poesio, M. (2000) “An empirically-based system for processing definite descriptions”. In: Computational Linguistics.
Vieira, R.; Salmon-Alt, S.; Schang, E. (2002) “Multilingual corpora annotation for processing definite descriptions”. In: Proceedings of the PorTAL 2002, Faro.
Vieira, R.; Gasperin, C.; Goulart, R.; Salmon-Alt, S. (2003) “From concrete to virtual annotation mark-up language: the case of COMMOn-REFs”. In Proceedings of the (ACL 2003) Workshop on Linguistic Annotation: Getting the Model Right, Sapporo.
Publicado
31/07/2004
Como Citar
COLLOVINI, Sandra; GOULART, Rodrigo; VIEIRA, Renata.
Identificação de Expressões Anafóricas e Não Anafóricas com Base na Estrutura do Sintagma. In: SIMPÓSIO BRASILEIRO DE TECNOLOGIA DA INFORMAÇÃO E DA LINGUAGEM HUMANA (STIL), 2. , 2004, Salvador/BA.
Anais [...].
Porto Alegre: Sociedade Brasileira de Computação,
2004
.
p. 104-113.
