14 papers
CrossAlpha: An Annual-Report Benchmark for Cross-Market Factor Research (with LLM Agents)
Qian Wang, Zhongyi Tong, Nuo Chen +2
Cross-market factor research studies whether firm-level signals from one or more markets can predict returns in a target market, but existing public benchmarks do not support cross…
EmoTrack: Robust Depression Tracking from Counseling Transcripts across Session Regimes
Zhaomin Wu, Jiayi Li, Bingsheng He
Text-based counseling is an important interface for AI mental-health support, where transcripts may be used to monitor depression severity and flag sessions requiring timely human…
LLM DNA: Tracing Model Evolution via Functional Representations
Zhaomin Wu, Haodong Zhao, Ziyang Wang +3
The explosive growth of large language models (LLMs) has created a vast but opaque landscape: millions of models exist, yet their evolutionary relationships through fine-tuning, di…
Beyond Prompt-Induced Lies: Investigating LLM Deception on Benign Prompts
Zhaomin Wu, Mingzhe Du, See-Kiong Ng +1
Large Language Models (LLMs) are widely deployed in reasoning, planning, and decision-making tasks, making their trustworthiness critical. A significant and underexplored risk is i…
WikiDBGraph: A Data Management Benchmark Suite for Collaborative Learning over Database Silos
Zhaomin Wu, Ziyang Wang, Bingsheng He
Relational databases are often fragmented across organizations, creating data silos that hinder distributed data management and mining. Collaborative learning (CL) -- techniques th…
ProtegoFed: Backdoor-Free Federated Instruction Tuning with Interspersed Poisoned Data
Haodong Zhao, Jinming Hu, Zhaomin Wu +7
Federated Instruction Tuning (FIT) enables collaborative instruction tuning of large language models across multiple organizations (clients) in a cross-silo setting without requiri…