output
20192026
most citedNTU RGB+D 120: A Large-Scale Benchmark for 3D Human Activity Understanding

1.8k citations

Showing 2023Show all

33 papers · 1 filter

cs.CV20236 cited

Continual Referring Expression Comprehension via Dual Modular Memorization

Heng Tao Shen, Cheng Chen, Peng Wang +3

Referring Expression Comprehension (REC) aims to localize an image region of a given object described by a natural-language expression. While promising performance has been demonst…

cs.CV202314 cited

Class Gradient Projection For Continual Learning

Cheng Chen, Ji Zhang, Jingkuan Song +1

Catastrophic forgetting is one of the most critical challenges in Continual Learning (CL). Recent approaches tackle this problem by projecting the gradient update orthogonal to the…

cs.CV20238 cited

Dynamic Compositional Graph Convolutional Network for Efficient Composite Human Motion Prediction

Wanying Zhang, Shen Zhao, Fanyang Meng +2

With potential applications in fields including intelligent surveillance and human-robot interaction, the human motion prediction task has become a hot research topic and also has…

eess.SY202314 cited

A Computationally Efficient Bi-level Coordination Framework for CAVs at Unsignalized Intersections

Jiping Luo, Tingting Zhang, Qinyu Zhang

In this paper, we investigate cooperative vehicle coordination for connected and automated vehicles (CAVs) at unsignalized intersections. To support high traffic throughput while r…

cs.MM20233 cited

Redistributing the Precision and Content in 3D-LUT-based Inverse Tone-mapping for HDR/WCG Display

Cheng Guo, Leidong Fan, Qian Zhang +3

ITM(inverse tone-mapping) converts SDR (standard dynamic range) footage to HDR/WCG (high dynamic range /wide color gamut) for media production. It happens not only when remastering…

cs.SD2023

Text-Only Domain Adaptation for End-to-End Speech Recognition through Down-Sampling Acoustic Representation

Jiaxu Zhu, Weinan Tong, Yaoxun Xu +6

Mapping two modalities, speech and text, into a shared representation space, is a research topic of using text-only data to improve end-to-end automatic speech recognition (ASR) pe…