5 papers
Motion-Adapter: A Diffusion Model Adapter for Text-to-Motion Generation of Compound Actions
Yue Jiang, Mingyu Yang, Liuyuxin Yang +3
Recent advances in generative motion synthesis have enabled the production of realistic human motions from diverse input modalities. However, synthesizing compound actions from tex…
AI-based experts' knowledge visualization of cultural heritage: A case study of Terracotta Warriors
Siyi Li, Yue Jiang, Bowen Jing +2
Advancements in 3D modeling,digital display technologies,and the growing availability of digital cultural heritage data have significantly improved the accuracy of heritage depicti…
One Shot Learning for Edge Detection on Point Clouds
Zhikun Tu, Yuhe Zhang, Yiou Jia +2
Each scanner possesses its unique characteristics and exhibits its distinct sampling error distribution. Training a network on a dataset that includes data collected from different…
RailVQA: A Benchmark and Framework for Efficient Interpretable Visual Cognition in Automatic Train Operation
Sen Zhang, Runmei Li, Shizhuang Deng +7
As Automatic Train Operation (ATO) advances toward GoA4 and beyond, it increasingly depends on efficient, reliable cab-view visual perception and decision-oriented inference to ens…
StoryTailor:A Zero-Shot Pipeline for Action-Rich Multi-Subject Visual Narratives
Jinghao Hu, Yuhe Zhang, GuoHua Geng +2
Generating multi-frame, action-rich visual narratives without fine-tuning faces a threefold tension: action text faithfulness, subject identity fidelity, and cross-frame background…