papers

Publications (12)

cs.CV2022

Learning Object-Language Alignments for Open-Vocabulary Object Detection

Chuang Lin, Peize Sun, Yi Jiang +5

Existing object detection methods are bounded in a fixed-set vocabulary by costly labeled data. When dealing with novel categories, the model has to be retrained with more bounding…

cs.CV2020

Multi-source Domain Adaptation for Visual Sentiment Classification

Chuang Lin, Sicheng Zhao, Lei Meng +1

Existing domain adaptation methods on visual sentiment classification typically are investigated under the single-source scenario, where the knowledge learned from a source domain…

cs.SI2014

Flow-based Influence Graph Visual Summarization

Lei Shi, Hanghang Tong, Jie Tang +1

Visually mining a large influence graph is appealing yet challenging. People are amazed by pictures of newscasting graph on Twitter, engaged by hidden citation networks in academic…

cs.CV2022

Multimodal Transformer with Variable-length Memory for Vision-and-Language Navigation

Chuang Lin, Yi Jiang, Jianfei Cai +3

Vision-and-Language Navigation (VLN) is a task that an agent is required to follow a language instruction to navigate to the goal position, which relies on the ongoing interactions…

cs.NI2014

Scale Congestion Control to Ultra-High Speed Ethernet

Wanchun Jiang, Fengyuan Ren, Xin Yue +1

Currently, Ethernet is broadly used in LAN, datacenter and enterprise networks, storage networks, high performance computing networks and so on. Along with the popularity of Ethern…

cs.CV2025

Waver: Wave Your Way to Lifelike Video Generation

Yifu Zhang, Hao Yang, Yuqi Zhang +7

We present Waver, a high-performance foundation model for unified image and video generation. Waver can directly generate videos with durations ranging from 5 to 10 seconds at a na…

cs.CV2024

Generative Region-Language Pretraining for Open-Ended Object Detection

Chuang Lin, Yi Jiang, Lizhen Qu +2

In recent research, significant attention has been devoted to the open-vocabulary object detection task, aiming to generalize beyond the limited number of classes labeled during tr…

cs.CV2024

Drive-1-to-3: Enriching Diffusion Priors for Novel View Synthesis of Real Vehicles

Chuang Lin, Bingbing Zhuang, Shanlin Sun +3

The recent advent of large-scale 3D data, e.g. Objaverse, has led to impressive progress in training pose-conditioned diffusion models for novel view synthesis. However, due to the…

cs.PF2014

Characterizing the Impact of the Workload on the Value of Dynamic Resizing in Data Centers

Kai Wang, Minghong Lin, Florin Ciucu +2

Energy consumption imposes a significant cost for data centers; yet much of that energy is used to maintain excess service capacity during periods of predictably low load. Resultan…

cs.HC2025

sEMG-Based Joint Angle Estimation via Hierarchical Spiking Attentional Feature Decomposition Network

Xin Zhou, Chuang Lin, Can Wang +1

Surface electromyography (sEMG) has demonstrated significant potential in simultaneous and proportional control (SPC). However, existing algorithms for predicting joint angles base…

cs.CV2020

Emotional Semantics-Preserved and Feature-Aligned CycleGAN for Visual Emotion Adaptation

Sicheng Zhao, Xuanbai Chen, Xiangyu Yue +7

Thanks to large-scale labeled training data, deep neural networks (DNNs) have obtained remarkable success in many vision and multimedia tasks. However, because of the presence of d…

cs.CV2018

Robust Multi-subspace Analysis Using Novel Column L0-norm Constrained Matrix Factorization

Binghui Wang, Chuang Lin

We study the underlying structure of data (approximately) generated from a union of independent subspaces. Traditional methods learn only one subspace, failing to discover the mult…