most citedCrisisViT: A Robust Vision Transformer for Crisis Image Classification

6 citations · 11 across the 5 of their papers we have counts for

collaborators

5 papers

cs.IR2024

CFIR: Fast and Effective Long-Text To Image Retrieval for Large Corpora

Zijun Long, Xuri Ge, Richard Mccreadie +1

Text-to-image retrieval aims to find the relevant images based on a text query, which is important in various use-cases, such as digital libraries, e-commerce, and multimedia datab…

cs.CV2024

Understanding and Mitigating Human-Labelling Errors in Supervised Contrastive Learning

Zijun Long, Lipeng Zhuang, George Killick +3

Human-annotated vision datasets inevitably contain a fraction of human mislabelled examples. While the detrimental effects of such mislabelling on supervised learning are well-rese…

cs.CV20246 cited

CrisisViT: A Robust Vision Transformer for Crisis Image Classification

Zijun Long, Richard McCreadie, Muhammad Imran

In times of emergency, crisis response agencies need to quickly and accurately assess the situation on the ground in order to deploy relevant services and resources. However, autho…

cs.IR20234 cited

Large Multi-modal Encoders for Recommendation

Zixuan Yi, Zijun Long, Iadh Ounis +2

In recent years, the rapid growth of online multimedia services, such as e-commerce platforms, has necessitated the development of personalised recommendation approaches that can e…

cs.CV20231 cited

When hard negative sampling meets supervised contrastive learning

Zijun Long, George Killick, Richard McCreadie +2

State-of-the-art image models predominantly follow a two-stage strategy: pre-training on large datasets and fine-tuning with cross-entropy loss. Many studies have shown that using…