3 papers
cs.CV2026
3D-PLOT-LLM: Part-Level Object Tokens for 3D Large Language Models
Jintang Xue, Xinyu Wang, Yixing Wu +2
3D multimodal large language models (3D MLLMs) describe a 3D object as a whole but cannot address, name, or reason about its parts. Prior part-aware attempts add segmentation decod…
cs.CV2024
Efficient Human-Object-Interaction (EHOI) Detection via Interaction Label Coding and Conditional Decision
Tsung-Shan Yang, Yun-Cheng Wang, Chengwei Wei +2
Human-Object Interaction (HOI) detection is a fundamental task in image understanding. While deep-learning-based HOI methods provide high performance in terms of mean Average Preci…
cs.CL2024
Word Embedding Dimension Reduction via Weakly-Supervised Feature Selection
Jintang Xue, Yun-Cheng Wang, Chengwei Wei +1
As a fundamental task in natural language processing, word embedding converts each word into a representation in a vector space. A challenge with word embedding is that as the voca…