most citedAutomatic annotation of multilingual text collections with a conceptual thesaurus

79 citations · 199 across the 8 of their papers we have counts for

collaborators

10 papers

cs.CL20069 cited

A tool set for the quick and efficient exploration of large document collections

Camelia Ignat, Bruno Pouliquen, Ralf Steinberger +1

We are presenting a set of multilingual text analysis tools that can help analysts in any field to explore large document collections quickly in order to determine whether the docu…

cs.CL20066 cited

Building and displaying name relations using automatic unsupervised analysis of newspaper articles

Bruno Pouliquen, Ralf Steinberger, Camelia Ignat +1

We present a tool that, from automatically recognised names, tries to infer inter-person relations in order to present associated people on maps. Based on an in-house Named Entity…

cs.CL2006

Geocoding multilingual texts: Recognition, disambiguation and visualisation

Bruno Pouliquen, Marco Kimler, Ralf Steinberger +8

We are presenting a method to recognise geographical references in free text. Our tool must work on various languages with a minimum of language-dependent resources, except a gazet…

cs.CL200635 cited

Exploiting multilingual nomenclatures and language-independent text features as an interlingua for cross-lingual text analysis applications

Ralf Steinberger, Bruno Pouliquen, Camelia Ignat

We are proposing a simple, but efficient basic approach for a number of multilingual and cross-lingual language technology applications that are not limited to the usual two or thr…

cs.CL200612 cited

Extending an Information Extraction tool set to Central and Eastern European languages

Camelia Ignat, Bruno Pouliquen, Antonio Ribeiro +1

In a highly multilingual and multicultural environment such as in the European Commission with soon over twenty official languages, there is an urgent need for text analysis tools…

cs.CL200653 cited

Automatic Identification of Document Translations in Large Multilingual Document Collections

Bruno Pouliquen, Ralf Steinberger, Camelia Ignat

Texts and their translations are a rich linguistic resource that can be used to train and test statistics-based Machine Translation systems and many other applications. In this pap…