5 papers
The Wikidata Query Logs Dataset
Sebastian Walter, Hannah Bast
We present the Wikidata Query Logs (WDQL) dataset, a dataset consisting of 335k question-query pairs over the Wikidata knowledge graph. It is over 11x larger than the largest exist…
GRISP: Guided Recurrent IRI Selection over SPARQL Skeletons
Sebastian Walter, Hannah Bast
We present GRISP (Guided Recurrent IRI Selection over SPARQL Skeletons), a novel SPARQL-based question-answering method over knowledge graphs using a fine-tuned small language mode…
Learning to Read Where to Look: Disease-Aware Vision-Language Pretraining for 3D CT
Simon Ging, Philipp Arnold, Sebastian Walter +6
Recent 3D CT vision-language models align volumes with reports via contrastive pretraining, but typically rely on limited public data and provide only coarse global supervision. We…
GRASP: Generic Reasoning And SPARQL Generation across Knowledge Graphs
Sebastian Walter, Hannah Bast
We propose a new approach for generating SPARQL queries on RDF knowledge graphs from natural language questions or keyword queries, using a large language model. Our approach does…
Using Knowledge Graphs to harvest datasets for efficient CLIP model training
Simon Ging, Sebastian Walter, Jelena BratuliÄ +3
Training high-quality CLIP models typically requires enormous datasets, which limits the development of domain-specific models -- especially in areas that even the largest CLIP mod…