works on

From the 1 of 11 linked papers with an AI index.

collaborators

11 papers

cs.CV2026

SiPhy: Single-Image Physical Property Reasoning

Hoang Le, Joonwoo Kwon, Elkhan Ismayilzada +2

Inferring physical properties such as mass, stiffness, and elasticity from a single image is essential for simulation and embodied AI, yet most existing approaches rely on multi-vi…

cs.CV2026

SceneBind: Binding What and Where Across Vision, Audio and Language

Mingfei Chen, Zijun Cui, Ruoke Zhang +2

SceneBind introduces an omni‑modal representation that jointly encodes what objects are and where they are in 3D space across vision, audio, and language, enabling cross‑modal scen…

cs.CV2026

Revisiting Euler-Angle Regression with Kolmogorov-Arnold Networks

Yangting Sun, Zijun Cui, Yufei Zhang

In many real-world systems, including articulated robots and biomechanical models, rotations are defined in joint space and naturally parameterized by Euler angles with bounded ran…

cs.LG2026

GEESE: Genotype-aware End-to-End Spatio-temporal Embedding for Behavioral Phenotyping

Yiran Ding, Yuen Gao, Chunqi Qian +1

Behavioral phenotyping of genetic animal models currently requires labor-intensive manual feature engineering that limits reproducibility and scalability. We present GEESE, an end-…

cs.SD2026

Do Joint Audio-Video Generation Models Understand Physics?

Zijun Cui, Xiulong Liu, Hao Fang +8

Joint audio-video generation models are rapidly approaching professional production quality, raising a central question: do they understand audio-visual physics, or merely generate…

cs.AI2026

Grading the Unspoken: Evaluating Tacit Reasoning in Quantum Field Theory and String Theory with LLMs

Xingyang Yu, Yinghuan Zhang, Yufei Zhang +1

Large language models have demonstrated impressive performance across many domains of mathematics and physics. One natural question is whether such models can support research in h…