Discriminating word senses with tourist walks in complex networks
arXiv:1306.3920 · doi:10.1140/epjb/e2013-40025-4
Abstract
Patterns of topological arrangement are widely used for both animal and human brains in the learning process. Nevertheless, automatic learning techniques frequently overlook these patterns. In this paper, we apply a learning technique based on the structural organization of the data in the attribute space to the problem of discriminating the senses of 10 polysemous words. Using two types of characterization of meanings, namely semantical and topological approaches, we have observed significative accuracy rates in identifying the suitable meanings in both techniques. Most importantly, we have found that the characterization based on the deterministic tourist walk improves the disambiguation process when one compares with the discrimination achieved with traditional complex networks measurements such as assortativity and clustering coefficient. To our knowledge, this is the first time that such deterministic walk has been applied to such a kind of problem. Therefore, our finding suggests that the tourist walk characterization may be useful in other related applications.
References in corpus (10)
- Characterization of complex networks: A survey of measurements
- Network properties of written human language
- Statistical keyword detection in literary corpora
- Consensus and ordering in language dynamics
- Combining Knowledge- and Corpus-based Word-Sense-Disambiguation Methods
- Structure-semantics interplay in complex networks and its effects on the predictability of similarity in texts
- On the use of topological features and hierarchical characterization for disambiguating names in collaborative networks
- Three-feature model to reproduce the topology of citation networks and the effects from authors' visibility on their h-index
- Identification of Literary Movements Using Complex Networks to Represent Texts
- Unveiling the relationship between complex networks metrics and word senses