2 papers
cs.LG2025
Fine-grained Token Allocation Via Operation Pruning for Efficient MLLMs
Aoming Liu, Reuben Tan, Boqing Gong +1
Token reduction accelerates Multimodal Large Language Models (MLLMs) by reducing excessive tokens, but overlooks structural redundancy differences, where critical and redundant mod…
cs.CV2025
Enhancing Virtual Try-On with Synthetic Pairs and Error-Aware Noise Scheduling
Nannan Li, Kevin J. Shih, Bryan A. Plummer
Given an isolated garment image in a canonical product view and a separate image of a person, the virtual try-on task aims to generate a new image of the person wearing the target…