51 papers
Handling covariate shift by model averaging
Yifan Zhang, Tianfa Xie, Xinyu Zhang
Distributional mismatch between the data used to construct a statistical procedure and the population to which it is ultimately applied is pervasive in modern data analysis. We stu…
TIEM: Temporal Integration of Hypergraph Evidence and Skill Memory for Event-Driven Financial Forecasting
Wenjin Liu, Shen Pang, Chenxi Wang +5
Event-driven catalyst-outcome forecasting increasingly uses retrieval- and memory-augmented large language model agents for prediction. However, training-data contamination and tem…
Predict, Then Retrieve: Cross-Instance Future-State Retrieval from Video Prefixes
Quynh Vo, Thong Nguyen, Vinh-Hien Do +2
We introduce Predictive State Retrieval (PSR), a task in which a model observes a short video prefix and a temporal question about an object's future state, then retrieves instance…
PowerAtlas: Towards Electricity-Computing Co-Scheduling for Power Systems
Kaiwen Jiang, Siya Xu, Ziyue Zhu +3
PowerAtlas is an LLM‑agent framework that jointly schedules electricity supply and computing workloads in data centers, ensuring grid operational constraints and computing service…
From Signals to Transfer: A Factorised Study of Probe-Based Uncertainty Estimation in Large Language Models
Ponhvoan Srey, Xiaobao Wu, Cong-Duy Nguyen +3
Probe-based uncertainty estimation (UE) has emerged as a prominent approach to detect hallucinations in Large Language Models (LLMs) by learning uncertainty from internal model sig…
EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments
Jundong Xu, Qingchuan Li, Jiaying Wu +11
Large language model (LLM) agents have achieved strong performance on a wide range of benchmarks, yet most evaluations assume static environments. In contrast, real-world deploymen…