collaborators

13 papers

cs.CV2025

UniHPR: Unified Human Pose Representation via Singular Value Contrastive Learning

Zhongyu Jiang, Wenhao Chai, Lei Li +3

In recent years, there has been a growing interest in developing effective alignment pipelines to generate unified representations from different modalities for multi-modal fusion…

cs.CV2025

Bayesian Optimization for Controlled Image Editing via LLMs

Chengkun Cai, Haoliang Liu, Xu Zhao +6

In the rapidly evolving field of image generation, achieving precise control over generated content and maintaining semantic consistency remain significant limitations, particularl…

cs.AI2025

The Role of Deductive and Inductive Reasoning in Large Language Models

Chengkun Cai, Xu Zhao, Haoliang Liu +5

Large Language Models (LLMs) have demonstrated impressive capabilities in reasoning tasks, yet their reliance on static prompt structures and limited adaptability to complex scenar…

cs.AI2025

Human Motion Instruction Tuning

Lei Li, Sen Jia, Jianhao Wang +6

This paper presents LLaMo (Large Language and Human Motion Assistant), a multimodal framework for human motion instruction tuning. In contrast to conventional instruction-tuning ap…

cs.CV2025

PackDiT: Joint Human Motion and Text Generation via Mutual Prompting

Zhongyu Jiang, Wenhao Chai, Zhuoran Zhou +3

Human motion generation has advanced markedly with the advent of diffusion models. Most recent studies have concentrated on generating motion sequences based on text prompts, commo…

cs.CV2025

MambaMOT: State-Space Model as Motion Predictor for Multi-Object Tracking

Hsiang-Wei Huang, Cheng-Yen Yang, Wenhao Chai +2

In the field of multi-object tracking (MOT), traditional methods often rely on the Kalman filter for motion prediction, leveraging its strengths in linear motion scenarios. However…