9 papers
FMMD: A multimodal open peer review dataset based on F1000Research
Zhenzhen Zhuang, Yuqing Fu, Jing Zhu +2
Automated scholarly paper review (ASPR) has entered the coexistence phase with traditional peer review, where artificial intelligence (AI) systems are increasingly incorporated int…
DeDPO: Debiased Direct Preference Optimization for Diffusion Models
Khiem Pham, Quang Nguyen, Tung Nguyen +4
Direct Preference Optimization (DPO) has emerged as a predominant alignment method for diffusion models, facilitating off-policy training without explicit reward modeling. However,…
FinDeepResearch: Evaluating Deep Research Agents in Rigorous Financial Analysis
Fengbin Zhu, Xiang Yao Ng, Ziyang Liu +19
Deep Research (DR) agents, powered by advanced Large Language Models (LLMs), have recently garnered increasing attention for their capability in conducting complex research tasks.…
LLM Optimization Unlocks Real-Time Pairwise Reranking
Jingyu Wu, Aditya Shrivastava, Jing Zhu +3
Efficiently reranking documents retrieved from information retrieval (IR) pipelines to enhance overall quality of Retrieval-Augmented Generation (RAG) system remains an important y…
Toward a Team of AI-made Scientists for Scientific Discovery from Gene Expression Data
Haoyang Liu, Yijiang Li, Jinglin Jian +7
Machine learning has emerged as a powerful tool for scientific discovery, enabling researchers to extract meaningful insights from complex datasets. For instance, it has facilitate…
HEAL: A Hypothesis-Based Preference-Aware Analysis Framework
Yifu Huo, Chenglong Wang, Qiren Zhu +5
Preference optimization methods like DPO have achieved remarkable performance in LLM alignment. However, the evaluation for these methods relies on a single response and overlooks…