9 papers
RecourseBench: A Modular Framework for Reproducible Algorithmic Recourse Evaluation
Zahra Khotanlou, Hashir Ahmed, Chenghao Tan +2
Algorithmic recourse methods provide counterfactual explanations that inform individuals of the actions required to overturn an unfavorable model decision. Despite rapid methodolog…
PaperFlow: Profiling, Recommending, and Adapting Across Daily Paper Streams
Fuqiang Wang, Song Tan, Zheng Guo +8
Scientific paper recommendation is typically evaluated as static ranking over a fixed candidate set, yet real scientific reading unfolds as a daily, longitudinal process in which i…
Thinking with Drafting: Optical Decompression via Logical Reconstruction
Jingxuan Wei, Honghao He, Caijun Jia +9
Existing multimodal large language models have achieved high-fidelity visual perception and exploratory visual generation. However, a precision paradox persists in complex reasonin…
Towards Robustness: A Critique of Current Vector Database Assessments
Zikai Wang, Qianxi Zhang, Baotong Lu +2
Vector databases are critical infrastructure in AI systems, and average recall is the dominant metric for their evaluation. Both users and researchers rely on it to choose and opti…
GGBench: A Geometric Generative Reasoning Benchmark for Unified Multimodal Models
Jingxuan Wei, Caijun Jia, Xi Bai +7
The advent of Unified Multimodal Models (UMMs) signals a paradigm shift in artificial intelligence, moving from passive perception to active, cross-modal generation. Despite their…
ResearchPulse: Building Method-Experiment Chains through Multi-Document Scientific Inference
Qi Chen, Jingxuan Wei, Zhuoya Yao +5
Understanding how scientific ideas evolve requires more than summarizing individual papers-it demands structured, cross-document reasoning over thematically related research. In th…