2 papers
cs.CV2026
ScalePredictor: Instance-aware Scale Learning for Accurate Quantization of Vision Transformers
Changjun Li, Runqing Jiang, Lian Xu +3
Vision Transformers have achieved remarkable success in many fields, yet their deployment on edge devices remains challenging due to their substantial computational demands. Post-T…
cs.CV2026
MLLMs Get It Right, Then Get It Wrong: Tracing and Correcting Late-Layer Textual Bias
Xingming Li, Ao Cheng, Qiyao Sun +4
When vision contradicts text, multimodal large language models (MLLMs) consistently favor text, even when images provide clear evidence otherwise. This bias poses risks for applica…