activity
20232026
most citedSequential Query Encoding For Complex Query Answering on Knowledge Graphs

5 citations · 11 across the 39 of their papers we have counts for

collaborators

41 papers

cs.LG2026

Same Outcome, Different Readout: What Does a Steerable Valence Direction in LLMs Represent?

Weihan Li, Xinlei Chen, Yuhan Song +2

Decodability and successful activation steering do not, by themselves, establish what an internal direction represents. This gap is especially consequential for welfare-relevant in…

cs.AI2026

AgentIdeaBench: Benchmarking Scientific Ideation in the Agent Era

Yunxiang Mo, Tianshi Zheng, Yisen Gao +7

Scientific ideation is the capacity to formulate novel and testable hypotheses from scientific evidence, and autonomous AI scientists depend on it. Existing evaluations largely ass…

cs.AI2026

MultivationBench: A Benchmark for Multimodal Sequential Motivation Reasoning

Kawai Chung, Chunkit Chan, Yauwai Yim +12

Multimodal Large Language Models have sparked significant interest due to their potential for social intelligence; however, their ability to perform sequential motivation reasoning…

cs.CV2026

Finding the Right Evidence: Factor-Guided Coarse-to-Fine Reasoning for Long Videos

Baixuan Xu, Yinyui Xu, Tianshi Zheng +9

While LVLMs rapidly improve, long-video question answering still remains challenging: relevant evidence is sparse, and question-relevant context often fails to provide cues that di…

cs.AI2026

Grading the Grader: Lessons from Evaluating an Agentic Data Analysis System

Tian Zheng, Kai-Tai Hsu

Agentic data analysis systems produce rich outputs, including code, numerical results, and verbal diagnostics. This makes them more challenging to evaluate than single-turn LLM res…

stat.ME2026

Target-Aware Linear Regression Under Distribution Shift

Zhewen Hou, Tian Zheng

Distribution shift between training and deployment is a pervasive challenge for modern AI systems. In many cases, the target marginals of covariates and response are known or speci…