4 papers
CLCE: An Approach to Refining Cross-Entropy and Contrastive Learning for Optimized Learning Fusion
Zijun Long, George Killick, Lipeng Zhuang +3
State-of-the-art pre-trained image models predominantly adopt a two-stage approach: initial unsupervised pre-training on large-scale datasets followed by task-specific fine-tuning…
Understanding and Mitigating Human-Labelling Errors in Supervised Contrastive Learning
Zijun Long, Lipeng Zhuang, George Killick +3
Human-annotated vision datasets inevitably contain a fraction of human mislabelled examples. While the detrimental effects of such mislabelling on supervised learning are well-rese…
RoboLLM: Robotic Vision Tasks Grounded on Multimodal Large Language Models
Zijun Long, George Killick, Richard McCreadie +1
Robotic vision applications often necessitate a wide range of visual perception tasks, such as object detection, segmentation, and identification. While there have been substantial…
MultiWay-Adapater: Adapting large-scale multi-modal models for scalable image-text retrieval
Zijun Long, George Killick, Richard McCreadie +1
As Multimodal Large Language Models (MLLMs) grow in size, adapting them to specialized tasks becomes increasingly challenging due to high computational and memory demands. Indeed,…