6 papers · 1 filter
Descrip3D: Enhancing Large Language Model-based 3D Scene Understanding with Object-Level Text Descriptions
Jintang Xue, Ganning Zhao, Jie-En Yao +5
Understanding 3D scenes goes beyond simply recognizing objects; it requires reasoning about the spatial and semantic relationships between them. Current 3D scene-language models of…
Green Video Camouflaged Object Detection
Xinyu Wang, Hong-Shuo Chen, Zhiruo Zhou +3
Camouflaged object detection (COD) aims to distinguish hidden objects embedded in an environment highly similar to the object. Conventional video-based COD (VCOD) methods explicitl…
Efficient Human-Object-Interaction (EHOI) Detection via Interaction Label Coding and Conditional Decision
Tsung-Shan Yang, Yun-Cheng Wang, Chengwei Wei +2
Human-Object Interaction (HOI) detection is a fundamental task in image understanding. While deep-learning-based HOI methods provide high performance in terms of mean Average Preci…
GreenCOD: A Green Camouflaged Object Detection Method
Hong-Shuo Chen, Yao Zhu, Suya You +2
We introduce GreenCOD, a green method for detecting camouflaged objects, distinct in its avoidance of backpropagation techniques. GreenCOD leverages gradient boosting and deep feat…
SemST: Semantically Consistent Multi-Scale Image Translation via Structure-Texture Alignment
Ganning Zhao, Wenhui Cui, Suya You +1
Unsupervised image-to-image (I2I) translation learns cross-domain image mapping that transfers input from the source domain to output in the target domain while preserving its sema…
Unsupervised Green Object Tracker (GOT) without Offline Pre-training
Zhiruo Zhou, Suya You, C. -C. Jay Kuo
Supervised trackers trained on labeled data dominate the single object tracking field for superior tracking accuracy. The labeling cost and the huge computational complexity hinder…