3 papers
cs.CV2025
GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning
V Team, Wenyi Hong, Wenmeng Yu +90
We present GLM-4.1V-Thinking, GLM-4.5V, and GLM-4.6V, a family of vision-language models (VLMs) designed to advance general-purpose multimodal understanding and reasoning. In this…
cs.IR2025
HCMRM: A High-Consistency Multimodal Relevance Model for Search Ads
Guobing Gan, Kaiming Gao, Li Wang +2
Search advertising is essential for merchants to reach the target users on short video platforms. Short video ads aligned with user search intents are displayed through relevance m…
cs.CL2022
MorphTE: Injecting Morphology in Tensorized Embeddings
Guobing Gan, Peng Zhang, Sunzhu Li +2
In the era of deep learning, word embeddings are essential when dealing with text tasks. However, storing and accessing these embeddings requires a large amount of space. This is n…