papers

Publications (34)

cs.LG2026

DRIVE: Distributional and Retrieval-Augmented Bidding with Value Evaluation

Miduo Cui, Haochen Wang, Shangqin Mao +6

Auto-bidding is a core component of real-time advertising systems, where decisions must optimize long-term performance under budget and cost constraints, while online exploration i…

cs.IR2026

Stream-aware Side Adaptation for Large Pre-trained Multimodal Embedding Models in Sequential Recommendation

Junchen Fu, Kaiwen Zheng, Ioannis Arapakis +4

Recently, large pretrained multimodal embedding models such as Qwen3-VL Embedding have shown strong promise for sequential recommendation, as they provide reusable semantic item re…

cs.CV2022

Factored Attention and Embedding for Unstructured-view Topic-related Ultrasound Report Generation

Fuhai Chen, Rongrong Ji, Chengpeng Dai +4

Echocardiography is widely used to clinical practice for diagnosis and treatment, e.g., on the common congenital heart defects. The traditional manual manipulation is error-prone d…

cs.CV2025

Multimodal Representation Learning Techniques for Comprehensive Facial State Analysis

Kaiwen Zheng, Xuri Ge, Junchen Fu +2

Multimodal foundation models have significantly improved feature representation by integrating information from multiple modalities, making them highly suitable for a broader set o…

cs.IR2024

R^3AG: First Workshop on Refined and Reliable Retrieval Augmented Generation

Zihan Wang, Xuri Ge, Joemon M. Jose +4

Retrieval-augmented generation (RAG) has gained wide attention as the key component to improve generative models with external knowledge augmentation from information retrieval. It…

cs.IR2024

CFIR: Fast and Effective Long-Text To Image Retrieval for Large Corpora

Zijun Long, Xuri Ge, Richard Mccreadie +1

Text-to-image retrieval aims to find the relevant images based on a text query, which is important in various use-cases, such as digital libraries, e-commerce, and multimedia datab…