7 papers
ChainClaw: A Layered Agent Framework for Reliable On-Chain Execution
Jiacheng Wei, Zhaoxin Fan, Xin Wen +5
General-purpose large language model agents have achieved strong performance on tool-augmented tasks, yet they rely on assumptions break down in blockchain environments. On-chain e…
Requirement--Evidence Alignment for Compositional E-Commerce Queries
Weihao Shen, Wei Chen, Fuwei Zhang +6
Compositional e-commerce queries express multiple requirements that must hold jointly, yet existing rerankers collapse these constraints into aggregate relevance and often promote…
Unpaired Modality-Agnostic Generative Recommendation
Weihao Shen, Wei Chen, Fuwei Zhang +6
Generative Recommendation (GR) formulates recommendation as autoregressive generation over discrete semantic identifiers (IDs). Although recent multimodal GR methods improve semant…
State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading
Yuanze Hu, Gen Li, Yuqin Lan +5
Multimodal large language models (MLLMs) have achieved impressive progress on general multimodal tasks, yet they remain brittle on dial-based measurement reading. In this paper, we…
Mosaic: Multimodal Jailbreak against Closed-Source VLMs via Multi-View Ensemble Optimization
Yuqin Lan, Gen Li, Yuanze Hu +6
Vision-Language Models (VLMs) are powerful but remain vulnerable to multimodal jailbreak attacks. Existing attacks mainly rely on either explicit visual prompt attacks or gradient-…
You Only Anonymize What Is Not Intent-Relevant: Suppressing Non-Intent Privacy Evidence
Weihao Shen, Yaxin Xu, Shuang Li +4
Anonymizing sensitive information in user text is essential for privacy, yet existing methods often apply uniform treatment across attributes, which can conflict with communicative…