The role of word sense disambiguation in automated text categorization

ABACUS/Manakin Repository

Show simple item record

dc.contributor.author Gómez Hidalgo, José María
dc.contributor.author De Buenaga Rodríguez, Manuel
dc.contributor.author Cortizo Pérez, José Carlos
dc.date.accessioned 2016-07-27T07:55:01Z
dc.date.available 2016-07-27T07:55:01Z
dc.date.issued 2005
dc.identifier.citation Gómez Hidalgo, J. M., De Buenaga Rodríguez, M., & Cortizo Pérez, J. C. (2005). The Role of word sense disambiguation in automated text categorization. Lecture Notes in Computer Science, 3513, 298-309. spa
dc.identifier.isbn 9783540260318
dc.identifier.issn 03029743
dc.identifier.uri http://hdl.handle.net/11268/5479
dc.description.abstract Automated Text Categorization has reached the levels of accuracy of human experts. Provided that enough training data is available, it is possible to learn accurate automatic classifiers by using Information Retrieval and Machine Learning Techniques. However, performance of this approach is damaged by the problems derived from language variation (specially polysemy and synonymy). We investigate how Word Sense Disambiguation can be used to alleviate these problems, by using two traditional methods for thesaurus usage in Information Retrieval, namely Query Expansion and Concept Indexing. These methods are evaluated on the problem of using the Lexical Database WordNet for text categorization, focusing on the Word Sense Disambiguation step involved. Our experiments demonstrate that rather simple dictionary methods, and baseline statistical approaches, can be used to disambiguate words and improve text representation and learning in both Query Expansion and Concept Indexing approaches. spa
dc.description.sponsorship SIN FINANCIACIÓN spa
dc.language.iso eng spa
dc.title The role of word sense disambiguation in automated text categorization spa
dc.type article spa
dc.description.impact 0.288 SJR (2005) Q2, 78/176 Computer science (miscellaneous); Q4, 75/101 Theoretical computer science spa
dc.identifier.doi 10.1007/11428817_27
dc.rights.accessRights closedAccess en
dc.subject.uem Inteligencia artificial - Aplicaciones spa
dc.subject.uem Lenguajes de ordenador spa
dc.subject.uem Lenguajes formales spa
dc.subject.unesco Lenguajes controlados spa
dc.subject.unesco Inteligencia artificial spa
dc.subject.unesco Robótica spa
dc.description.filiation UEM spa
dc.peerreviewed Si spa

Files in this item

Files Size Format View

There are no files associated with this item.

This item appears in the following Collection(s)

Show simple item record