most citedArchiveSpark: Efficient Web Archive Access, Extraction and Derivation

25 citations · 96 across the 11 of their papers we have counts for

collaborators

11 papers

cs.CL20171 cited

Named Entity Evolution Recognition on the Blogosphere

Helge Holzmann, Nina Tahmasebi, Thomas Risse

Advancements in technology and culture lead to changes in our language. These changes create a gap between the language known by users and the language stored in digital archives.…

cs.CL2017

Extraction of Evolution Descriptions from the Web

Helge Holzmann, Thomas Risse

The evolution of named entities affects exploration and retrieval tasks in digital libraries. An information retrieval system that is aware of name changes can actively support use…

cs.CL20176 cited

Named Entity Evolution Analysis on Wikipedia

Helge Holzmann, Thomas Risse

Accessing Web archives raises a number of issues caused by their temporal characteristics. Additional knowledge is needed to find and understand older texts. Especially entities me…

cs.CL2017

Insights into Entity Name Evolution on Wikipedia

Helge Holzmann, Thomas Risse

Working with Web archives raises a number of issues caused by their temporal characteristics. Depending on the age of the content, additional knowledge might be needed to find and…

cs.DL20171 cited

Linking Mathematical Software in Web Archives

Helge Holzmann, Mila Runnwerth, Wolfram Sperber

The Web is our primary source of all kinds of information today. This includes information about software as well as associated materials, like source code, documentation, related…

cs.DL20178 cited

Archiving Software Surrogates on the Web for Future Reference

Helge Holzmann, Wolfram Sperber, Mila Runnwerth

Software has long been established as an essential aspect of the scientific process in mathematics and other disciplines. However, reliably referencing software in scientific publi…