3 papers
cs.CL2025
From scratch to silver: Creating trustworthy training data for patent-SDG classification using Large Language Models
Grazia Sveva Ascione, Nicolò Tamagnone
Classifying patents by their relevance to the UN Sustainable Development Goals (SDGs) is crucial for tracking how innovation addresses global challenges. However, the absence of a…
cs.CL2024
A comparative analysis of embedding models for patent similarity
Grazia Sveva Ascione, Valerio Sterzi
This paper makes two contributions to the field of text-based patent similarity. First, it compares the performance of different kinds of patent-specific pretrained embedding model…
cs.IR2024
Presenting Terrorizer: an algorithm for consolidating company names in patent assignees
Grazia Sveva Ascione, Valerio Sterzi
The problem of disambiguation of company names poses a significant challenge in extracting useful information from patents. This issue biases research outcomes as it mostly underes…