output
20022008
most citedClustered Multi-Task Learning: A Convex Formulation

285 citations

Showing cs.IRShow all

5 papers · 1 filter

cs.IR2007

Entity Ranking in Wikipedia

Anne-Marie Vercoustre, James A. Thom, Jovan Pehcevski

The traditional entity extraction problem lies in the ability of extracting named entities from plain text using natural language processing techniques and intensive training from…

cs.IR200713 cited

Use of Wikipedia Categories in Entity Ranking

James A. Thom, Jovan Pehcevski, Anne-Marie Vercoustre

Wikipedia is a useful source of knowledge that has many applications in language processing and knowledge representation. The Wikipedia category graph can be compared with the clas…

cs.IR20055 cited

Experiments in Clustering Homogeneous XML Documents to Validate an Existing Typology

Thierry Despeyroux, Yves Lechevallier, Brigitte Trousse +1

This paper presents some experiments in clustering homogeneous XMLdocuments to validate an existing classification or more generally anorganisational structure. Our approach integr…

cs.IR20055 cited

Enhancing Content-And-Structure Information Retrieval using a Native XML Database

Jovan Pehcevski, James A. Thom, Anne-Marie Vercoustre

Three approaches to content-and-structure XML retrieval are analysed in this paper: first by using Zettair, a full-text information retrieval system; second by using eXist, a nativ…

cs.IR20059 cited

Hybrid XML Retrieval: Combining Information Retrieval and a Native XML Database

Jovan Pehcevski, James A. Thom, Anne-Marie Vercoustre

This paper investigates the impact of three approaches to XML retrieval: using Zettair, a full-text information retrieval system; using eXist, a native XML database; and using a hy…