5 papers
RA-CAD: Learning Post-Execution Critique for State-Aware Text-to-CAD Generation
Shuhao Yan, Changhao He, Xi Peng +1
Text-to-CAD generation translates natural-language design intent into editable and executable parametric computer-aided design (CAD) codes, reducing the expertise and effort requir…
Active-SWE: Benchmarking Coding Agents for Proactive Bug Fixing without Issue Reports
Haobin Li, Ping Deng, Weizhong Qian +4
Coding agents powered by large language models (LLMs) are increasingly adopted in software engineering (SWE) scenarios, capable of fixing a specific bug in large-scale codebase. Ho…
Robust Multi-view Clustering against Imperfect Information
Zhichao Huang, Haochen Zhou, Hao Wang +2
Real-world multi-view data always suffer from imperfect information problem, where the view-specific observations are absent (i.e., Incomplete Views, IV) and cross-view corresponde…
ARK: A Dual-Axis Multimodal Retrieval Benchmark along Reasoning and Knowledge
Yijie Lin, Guofeng Ding, Haochen Zhou +3
Existing multimodal retrieval benchmarks largely emphasize semantic matching on daily-life images and offer limited diagnostics of professional knowledge and complex reasoning. To…
Toward Robust and Harmonious Adaptation for Cross-modal Retrieval
Haobin Li, Mouxing Yang, Xi Peng
Recently, the general-to-customized paradigm has emerged as the dominant approach for Cross-Modal Retrieval (CMR), which reconciles the distribution shift problem between the sourc…