2 papers
cs.IR2026
UniNote: A Unified Embedding Model for Multimodal Representation and Ranking
Jinghan Zhao, Wenwei Jin, Anqi Li +5
Item-to-Item (I2I) retrieval is a fundamental part of modern content platforms, supporting critical industrial workflows from recommendation engines to content auditing. While mult…
cs.CV2025
Learning Procedural-aware Video Representations through State-Grounded Hierarchy Unfolding
Jinghan Zhao, Yifei Huang, Feng Lu
Learning procedural-aware video representations is a key step towards building agents that can reason about and execute complex tasks. Existing methods typically address this probl…