7 citations · 7 across the 2 of their papers we have counts for
1 paper · 1 filter
Zilin Xiao, Ming Gong, Paola Cascante-Bonilla +3
We introduce AutoVER, an Autoregressive model for Visual Entity Recognition. Our model extends an autoregressive Multi-modal Large Language Model by employing retrieval augmented c…