Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models
Saeed Ranjbar Alvar, Gursimran Singh, Mohammad Akbari +1
Large Multimodal Models (LMMs) have emerged as powerful models capable of understanding various data modalities, including text, images, and videos. LMMs encode both text and visua…
cs.CV2024
Towards Secure and Usable 3D Assets: A Novel Framework for Automatic Visible Watermarking
Gursimran Singh, Tianxi Hu, Mohammad Akbari +2
3D models, particularly AI-generated ones, have witnessed a recent surge across various industries such as entertainment. Hence, there is an alarming need to protect the intellectu…