5 papers
Hybrid Advantage Estimation with Unified Critic for VLM Agentic Reinforcement Learning
Wenxuan Zhang, Yuhui Wang, Donggang Jia +5
Large Vision-Language Models (VLMs) now act as agents in interactive environments, where success requires coherent reasoning and decision-making across turns. Although end-to-end t…
Chat Modeling: Interaction-Enhanced Agent Framework for Visualizing Literature-Grounded Biological Structures
Donggang Jia, Yunhai Wang, Ivan Viola
Bioscientists frequently seek to visualize the biological systems they have empirically characterized and reported in the literature. Realizing such visualizations requires biologi…
AIvaluateXR: An Evaluation Framework for on-Device AI in XR with Benchmarking Results
Dawar Khan, Xinyu Liu, Omar Mena +3
The deployment of large language models (LLMs) on extended reality (XR) devices has great potential to advance the field of human-AI interaction. In the case of direct, on-device m…
SkinningGS: Editable Dynamic Human Scene Reconstruction Using Gaussian Splatting Based on a Skinning Model
Da Li, Donggang Jia, Markus Hadwiger +1
Reconstructing an interactive human avatar and the background from a monocular video of a dynamic human scene is highly challenging. In this work we adopt a strategy of point cloud…
RaRa Clipper: A Clipper for Gaussian Splatting Based on Ray Tracer and Rasterizer
Da Li, Donggang Jia, Yousef Rajeh +2
With the advancement of Gaussian Splatting techniques, a growing number of datasets based on this representation have been developed. However, performing accurate and efficient cli…