2 papers
cs.CV2025
IAP: Improving Continual Learning of Vision-Language Models via Instance-Aware Prompting
Hao Fu, Hanbin Zhao, Jiahua Dong +3
Recent pre-trained vision-language models (PT-VLMs) often face a Multi-Domain Task Incremental Learning (MTIL) scenario in practice, where several classes and domains of multi-moda…
cs.CV2025
TextToucher: Fine-Grained Text-to-Touch Generation
Jiahang Tu, Hao Fu, Fengyu Yang +3
Tactile sensation plays a crucial role in the development of multi-modal large models and embodied intelligence. To collect tactile data with minimal cost as possible, a series of…