18 papers
Reachability Is Not Realization: Tracing the Sources of LLM Benchmark Gains
Yanchao Li, Wanhao Liu, Jiaqing Xie +4
Benchmark gains are often treated as evidence of greater LLM capability. Yet the same gain can reflect different changes in model behavior. A model may reach new answers, or produc…
Disagree to Accelerate: Closing the Loop on Diffusion Feature Forecasts
Yanchao Li, Jiaqing Xie, Ben Gao +6
Training-free feature forecasting accelerates diffusion sampling by predicting features at skipped denoising steps. Recent work has mainly focused on designing stronger forecasters…
ABOPD: Antibody CDR Design via On-Policy Distillation
Zhuo Yang, Jiaying He, Jiaqing Xie +5
Antibodies are essential therapeutic molecules, and their complementarity-determining regions (CDRs) form the primary antigen-recognition interface. Recent protein generative model…
Eigenbasis-Independent Learnable Spectral Positional Encodings for Directed Graphs via Hermitian Block Krylov Subspaces
Jiaqing Xie, Yuxin Wang
Spectral positional encodings (PEs) for \emph{directed} graphs face two obstacles: magnetic Laplacians require an Hermitian eigendecomposition per potential, and their com…
OmniMatBench: A Human-Calibrated Multimodal Reasoning Benchmark Across 19 Materials Science Subfields
Wanhao Liu, Jiaqing Xie, Qian Tan +10
As multimodal language models play an increasingly important role in scientific research, materials science offers a critical testbed due to its interdisciplinary, multimodal, and…
SkillsInjector: Dynamic Skill Context Construction for LLM Agents
Yanchao Li, Wanhao Liu, Ben Gao +5
LLM agents now draw on growing skill libraries to handle complex tasks. However, injecting more skills does not always improve task completion and can even degrade it. Existing met…