collaborators

7 papers

cs.CV2026

Teaching Vision-Language Models to Use the Scale They Are Given: Label-Free Equivariance Training for Metric Physical Reasoning

Kaizhen Tan, Yang Feng, Heqing Du +3

Metric questions about video require vision-language models to use supplied real-world references to convert visual measurements into physical units. Yet we find that current model…

cs.AI2026

Reasoning Shortcuts and Value Symmetries: What Symmetry Permits, Architecture Realizes, and Optimization Selects

Xin Xu

Reasoning shortcuts are rule solutions that reach correct predictions through unintended concepts. A recent framework of Takemura, Inoue, and Nishino analyzes them through an autom…

cs.LG2026

Beyond Multimodal Alignment: Certifying Physical Language through Response Substitution and Ordered Execution

Kaizhen Tan, Xin Xu, Siru Tao +4

World models increasingly treat compact multimodal representations as interfaces between perception and physical interaction, yet existing probes do not establish whether different…

cs.CR2026

Renaming or Tightness: Enforcing Disjunctive Information Flow Policies

Xin Xu, Siru Tao, Kaizhen Tan

A disjunctive policy allows a value to depend on at most one of two secrets and never on both: an analyst may consult one client's file or the other's, a share of a split secret ma…

cs.LG2026

What Can Latent World Models Know? Physical Parameter Identifiability in Multimodal Predictive Representations

Kaizhen Tan, Xin Xu, Siru Tao +4

A central premise of latent world models is that predicting the future forces a representation to internalize the physics of its environment. Which physical quantities does a train…

cs.GT2026

Collusion with Competitive Marginals: Price-Level Audits Are Blind by Construction

Xin Xu, Chengrui Wu, Jiayu Lu +3

Empirical work on algorithmic collusion asks one question of the data: are prices supracompetitive? We show this can be answered "no" by a conspiracy that is nonetheless profitable…