56 citations · 173 across the 15 of their papers we have counts for
7 papers · 1 filter
IterNet: Retinal Image Segmentation Utilizing Structural Redundancy in Vessel Networks
Liangzhi Li, Manisha Verma, Yuta Nakashima +2
Retinal vessel segmentation is of great interest for diagnosis of retinal vascular diseases. To further improve the performance of vessel segmentation, we propose IterNet, a new mo…
Historical and Modern Features for Buddha Statue Classification
Benjamin Renoust, Matheus Oliveira Franca, Jacob Chan +7
While Buddhism has spread along the Silk Roads, many pieces of art have been displaced. Only a few experts may identify these works, subjectively to their experience. The construct…
KnowIT VQA: Answering Knowledge-Based Questions about Videos
Noa Garcia, Mayu Otani, Chenhui Chu +1
We propose a novel video understanding task by fusing knowledge-based and video question answering. First, we introduce KnowIT VQA, a video dataset with 24,282 human-generated ques…
BUDA.ART: A Multimodal Content-Based Analysis and Retrieval System for Buddha Statues
Benjamin Renoust, Matheus Oliveira Franca, Jacob Chan +6
We introduce BUDA.ART, a system designed to assist researchers in Art History, to explore and analyze an archive of pictures of Buddha statues. The system combines different CBIR a…
Understanding Art through Multi-Modal Retrieval in Paintings
Noa Garcia, Benjamin Renoust, Yuta Nakashima
In computer vision, visual arts are often studied from a purely aesthetics perspective, mostly by analysing the visual appearance of an artistic reproduction to infer its style, it…
Rethinking the Evaluation of Video Summaries
Mayu Otani, Yuta Nakashima, Esa Rahtu +1
Video summarization is a technique to create a short skim of the original video while preserving the main stories/content. There exists a substantial interest in automatizing this…