3 papers
cs.IR2026
Interactive Multi-Turn Retrieval for Health Videos
Chengzheng Wu, Ke Qiu, Baoming Zhang +3
The growing availability of health-related instructional videos creates new opportunities for clinical training, patient rehabilitation, and health education, yet existing retrieva…
cs.CV2025
TernaryCLIP: Efficiently Compressing Vision-Language Models with Ternary Weights and Distilled Knowledge
Shu-Hao Zhang, Wei-Cheng Tang, Chen Wu +5
Recent years have witnessed an increasing interest in image-text contrastive modeling, exemplified by models such as Contrastive Language-Image Pretraining (CLIP). In this paper, w…
cs.CL2024
Efficient Ternary Weight Embedding Model: Bridging Scalability and Performance
Jiayi Chen, Chen Wu, Shaoqun Zhang +3
Embedding models have become essential tools in both natural language processing and computer vision, enabling efficient semantic search, recommendation, clustering, and more. Howe…