1 citations · 2 across the 3 of their papers we have counts for
3 papers · 1 filter
Language as the Medium: Multimodal Video Classification through text only
Laura Hanu, Anita L. Verő, James Thewlis
Despite an exciting new wave of multimodal machine learning models, current approaches still struggle to interpret the complex contextual relationships between the different modali…
VTC: Improving Video-Text Retrieval with User Comments
Laura Hanu, James Thewlis, Yuki M. Asano +1
Multi-modal retrieval is an important problem for many applications, such as recommendation and search. Current benchmarks and even datasets are often manually constructed and cons…
Deep Industrial Espionage
Samuel Albanie, James Thewlis, Sebastien Ehrhardt +1
The theory of deep learning is now considered largely solved, and is well understood by researchers and influencers alike. To maintain our relevance, we therefore seek to apply our…