most citedEvaluating Reasoning LLMs for Suicide Screening with the Columbia-Suicide Severity Rating Scale

1 citations · 1 across the 7 of their papers we have counts for

collaborators

8 papers

cs.CV2026

Teaching Vision-Language Models to Use the Scale They Are Given: Label-Free Equivariance Training for Metric Physical Reasoning

Kaizhen Tan, Yang Feng, Heqing Du +3

Metric questions about video require vision-language models to use supplied real-world references to convert visual measurements into physical units. Yet we find that current model…

cs.LG2026

Beyond Multimodal Alignment: Certifying Physical Language through Response Substitution and Ordered Execution

Kaizhen Tan, Xin Xu, Siru Tao +4

World models increasingly treat compact multimodal representations as interfaces between perception and physical interaction, yet existing probes do not establish whether different…

cs.CR2026

Renaming or Tightness: Enforcing Disjunctive Information Flow Policies

Xin Xu, Siru Tao, Kaizhen Tan

A disjunctive policy allows a value to depend on at most one of two secrets and never on both: an analyst may consult one client's file or the other's, a share of a split secret ma…

cs.LG2026

What Can Latent World Models Know? Physical Parameter Identifiability in Multimodal Predictive Representations

Kaizhen Tan, Xin Xu, Siru Tao +4

A central premise of latent world models is that predicting the future forces a representation to internalize the physics of its environment. Which physical quantities does a train…

cs.GT2026

Collusion with Competitive Marginals: Price-Level Audits Are Blind by Construction

Xin Xu, Chengrui Wu, Jiayu Lu +3

Empirical work on algorithmic collusion asks one question of the data: are prices supracompetitive? We show this can be answered "no" by a conspiracy that is nonetheless profitable…

cs.LG2026

Automorphism-Induced Non-Canonicity in Top-k Explanations of Graph Neural Networks

Xin Xu, Siru Tao, Kaizhen Tan

A gradient-based GNN explainer given a molecule with two chemically equivalent nitro groups assigns them attribution scores that are equal to the last bit. It cannot do otherwise:…