3 papers
cs.CV2026
FireRed-OCR Technical Report
Hao Wu, Haoran Lou, Xinyue Li +19
We present FireRed-OCR, a systematic framework to specialize general VLMs into high-performance OCR models. Large Vision-Language Models (VLMs) have demonstrated impressive general…
cs.IR2025
HyMiRec: A Hybrid Multi-interest Learning Framework for LLM-based Sequential Recommendation
Jingyi Zhou, Cheng Chen, Kai Zuo +5
Large language models (LLMs) have recently demonstrated strong potential for sequential recommendation. However, current LLM-based approaches face critical limitations in modeling…
cs.IR2025
Cross-Scenario Unified Modeling of User Interests at Billion Scale
Manjie Xu, Cheng Chen, Xin Jia +9
User interests on content platforms are inherently diverse, manifesting through complex behavioral patterns across heterogeneous scenarios such as search, feed browsing, and conten…