6 papers
Evaluating Newtonian Mechanics in Video Generative Models with Real Physical Systems
Antonios Tragoudaras, Chenyu Zhang, Daniil Cherniavskii +7
Recent advances in image and video generation raise hopes that these models possess world modeling capabilities-the ability to generate realistic, physically plausible videos. This…
Relational Anatomical Supervision for Accurate 3D Multi-Chamber Cardiac Mesh Reconstruction
Chenyu Zhang, Yihao Luo, Lei Zhu +3
Accurate reconstruction of multi-chamber cardiac anatomy from medical images is a cornerstone for patient-specific modeling, physiological simulation, and interventional planning.…
When One Modality Sabotages the Others: A Diagnostic Lens on Multimodal Reasoning
Chenyu Zhang, Minsol Kim, Shohreh Ghorbani +4
Despite rapid growth in multimodal large language models (MLLMs), their reasoning traces remain opaque: it is often unclear which modality drives a prediction, how conflicts are re…
Ensembling Large Language Models to Characterize Affective Dynamics in Student-AI Tutor Dialogues
Chenyu Zhang, Sharifa Alghowinem, Cynthia Breazeal
While recent studies have examined the leaning impact of large language model (LLM) in educational contexts, the affective dynamics of LLM-mediated tutoring remain insufficiently u…
Spreading Depolarization Detection in Electrocorticogram Spectrogram Imaging by Deep Learning: Is It Just About Delta Band?
Jeanne Boyer-Chammard, Yinzhe Wu, Chenyu Zhang +4
Prevention of secondary brain injury is a core aim of neurocritical care, with Spreading Depolarizations (SDs) recognized as a significant independent cause. SDs are typically moni…
Explanation for Trajectory Planning using Multi-modal Large Language Model for Autonomous Driving
Shota Yamazaki, Chenyu Zhang, Takuya Nanri +5
End-to-end style autonomous driving models have been developed recently. These models lack interpretability of decision-making process from perception to control of the ego vehicle…