The role of word sense disambiguation in automated text categorization

dc.contributor.authorGómez Hidalgo, José María
dc.contributor.authorDe Buenaga Rodríguez, Manuel
dc.contributor.authorCortizo Pérez, José Carlos
dc.date.accessioned2016-07-27T07:55:01Z
dc.date.available2016-07-27T07:55:01Z
dc.date.issued2005
dc.description.abstractAutomated Text Categorization has reached the levels of accuracy of human experts. Provided that enough training data is available, it is possible to learn accurate automatic classifiers by using Information Retrieval and Machine Learning Techniques. However, performance of this approach is damaged by the problems derived from language variation (specially polysemy and synonymy). We investigate how Word Sense Disambiguation can be used to alleviate these problems, by using two traditional methods for thesaurus usage in Information Retrieval, namely Query Expansion and Concept Indexing. These methods are evaluated on the problem of using the Lexical Database WordNet for text categorization, focusing on the Word Sense Disambiguation step involved. Our experiments demonstrate that rather simple dictionary methods, and baseline statistical approaches, can be used to disambiguate words and improve text representation and learning in both Query Expansion and Concept Indexing approaches.spa
dc.description.filiationUEMspa
dc.description.impact0.288 SJR (2005) Q2, 78/176 Computer science (miscellaneous); Q4, 75/101 Theoretical computer sciencespa
dc.description.sponsorshipSin financiaciónspa
dc.identifier.citationGómez Hidalgo, J. M., De Buenaga Rodríguez, M., & Cortizo Pérez, J. C. (2005). The Role of word sense disambiguation in automated text categorization. Lecture Notes in Computer Science, 3513, 298-309.spa
dc.identifier.doi10.1007/11428817_27
dc.identifier.isbn9783540260318
dc.identifier.issn03029743
dc.identifier.urihttp://hdl.handle.net/11268/5479
dc.language.isoengspa
dc.peerreviewedSispa
dc.rights.accessRightsrestricted accessen
dc.subject.uemInteligencia artificial - Aplicacionesspa
dc.subject.uemLenguajes de ordenadorspa
dc.subject.uemLenguajes formalesspa
dc.subject.unescoLenguajes controladosspa
dc.subject.unescoInteligencia artificialspa
dc.subject.unescoRobóticaspa
dc.titleThe role of word sense disambiguation in automated text categorizationspa
dc.typejournal articlespa
dspace.entity.typePublication
relation.isAuthorOfPublication76a395e8-090d-4187-9a3c-420063e1f44f
relation.isAuthorOfPublicatione1ae5b27-3248-41df-ac24-a38ed621e0f9
relation.isAuthorOfPublication.latestForDiscovery76a395e8-090d-4187-9a3c-420063e1f44f

Files