activity
20232025
collaborators
Showing cs.CVShow all

6 papers · 1 filter

cs.CV2025

Descrip3D: Enhancing Large Language Model-based 3D Scene Understanding with Object-Level Text Descriptions

Jintang Xue, Ganning Zhao, Jie-En Yao +5

Understanding 3D scenes goes beyond simply recognizing objects; it requires reasoning about the spatial and semantic relationships between them. Current 3D scene-language models of…

cs.CV2025

Green Video Camouflaged Object Detection

Xinyu Wang, Hong-Shuo Chen, Zhiruo Zhou +3

Camouflaged object detection (COD) aims to distinguish hidden objects embedded in an environment highly similar to the object. Conventional video-based COD (VCOD) methods explicitl…

cs.CV20242 cited

Efficient Human-Object-Interaction (EHOI) Detection via Interaction Label Coding and Conditional Decision

Tsung-Shan Yang, Yun-Cheng Wang, Chengwei Wei +2

Human-Object Interaction (HOI) detection is a fundamental task in image understanding. While deep-learning-based HOI methods provide high performance in terms of mean Average Preci…

cs.CV2024

GreenCOD: A Green Camouflaged Object Detection Method

Hong-Shuo Chen, Yao Zhu, Suya You +2

We introduce GreenCOD, a green method for detecting camouflaged objects, distinct in its avoidance of backpropagation techniques. GreenCOD leverages gradient boosting and deep feat…

cs.CV2023

SemST: Semantically Consistent Multi-Scale Image Translation via Structure-Texture Alignment

Ganning Zhao, Wenhui Cui, Suya You +1

Unsupervised image-to-image (I2I) translation learns cross-domain image mapping that transfers input from the source domain to output in the target domain while preserving its sema…

cs.CV2023

Unsupervised Green Object Tracker (GOT) without Offline Pre-training

Zhiruo Zhou, Suya You, C. -C. Jay Kuo

Supervised trackers trained on labeled data dominate the single object tracking field for superior tracking accuracy. The labeling cost and the huge computational complexity hinder…