1 paper
Zhanpeng Chen, Mingxiao Li, Ziyang Chen +3
Vision-language Models (VLMs) have shown remarkable capabilities in advancing general artificial intelligence, yet the irrational encoding of visual positions persists in inhibitin…