Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Look Less, Think Faster: Joint Token-Compute Adaptation for Multimodal LLMs
Pengcheng Wang, Zhiquan Wang, Jayoung Lee +5
Multimodal Large Language Models (MLLMs) have recently demonstrated strong performance across vision-language tasks. However, their high inference cost, arising from both the large…
cs.CV2020
ApproxDet: Content and Contention-Aware Approximate Object Detection for Mobiles
Ran Xu, Chen-lin Zhang, Pengcheng Wang +5
Advanced video analytic systems, including scene classification and object detection, have seen widespread success in various domains such as smart cities and autonomous transporta…