2 papers
cs.LG2026
Metadata-Free Meta-Reweighted Direct Preference Optimization under Noisy Preference Labels
Hua Qu, Yifan Li, Xiaodong Yuan
Direct Preference Optimization (DPO) has become an important method for aligning large language models (LLMs) with human preferences because it removes the need for explicit reward…
cs.CL2026
FinInvest-GTCN: Explainable Graph-Temporal-Causal Modeling for Risk-Aware Investment Decision Optimization
Junyan Tan, Yifan Li, Minghao Wang +2
Venture capital (VC) investment decisions face distinct challenges, such as multi-source heterogeneous data, non-stationary time series, and the demand for explainable predictions…