6 citations · 6 across the 4 of their papers we have counts for
4 papers · 1 filter
Aligned with LLM: a new multi-modal training paradigm for encoding fMRI activity in visual cortex
Shuxiao Ma, Linyuan Wang, Senbao Hou +1
Recently, there has been a surge in the popularity of pre trained large language models (LLMs) (such as GPT-4), sweeping across the entire Natural Language Processing (NLP) and Com…
A Multimodal Visual Encoding Model Aided by Introducing Verbal Semantic Information
Shuxiao Ma, Linyuan Wang, Bin Yan
Biological research has revealed that the verbal semantic information in the brain cortex, as an additional source, participates in nonverbal semantic tasks, such as visual encodin…
Exploring Transformers for Open-world Instance Segmentation
Jiannan Wu, Yi Jiang, Bin Yan +3
Open-world instance segmentation is a rising task, which aims to segment all objects in the image by learning from a limited number of base-category objects. This task is challengi…
Towards Grand Unification of Object Tracking
Bin Yan, Yi Jiang, Peize Sun +4
We present a unified method, termed Unicorn, that can simultaneously solve four tracking problems (SOT, MOT, VOS, MOTS) with a single network using the same model parameters. Due t…