2 papers
cs.CV2026
Multi-Perspective Subimage CLIP with Keyword Guidance for Remote Sensing Image-Text Retrieval
Yifan Li, Shiying Wang, Jianqiang Huang
Vision-Language Pre-training (VLP) models like CLIP have significantly advanced Remote Sensing Image-Text Retrieval (RSITR). However, existing methods predominantly rely on coarse-…
cs.HC2024
Evolving Agents: Interactive Simulation of Dynamic and Diverse Human Personalities
Jiale Li, Jiayang Li, Jiahao Chen +5
Human-like Agents with diverse and dynamic personalities could serve as an essential design probe in the process of user-centered design, thereby enabling designers to enhance the…