4 papers
Transferable Delay-Aware Reinforcement Learning via Implicit Causal Graph Modeling
Chenran Zhao, Dianxi Shi, Yaowen Zhang +2
Random delays weaken the temporal correspondence between actions and subsequent state feedback, making it difficult for agents to identify the true propagation process of action ef…
Delay-Empowered Causal Hierarchical Reinforcement Learning
Chenran Zhao, Dianxi Shi, Haotian Wang +4
Many real-world tasks involve delayed effects, where the outcomes of actions emerge after varying time lags. Existing delay-aware reinforcement learning methods often rely on state…
Improved Evidence Extraction and Metrics for Document Inconsistency Detection with LLMs
Nelvin Tan, Yaowen Zhang, James Asikin Cheung +3
Large language models (LLMs) are becoming useful in many domains due to their impressive abilities that arise from large training datasets and large model sizes. However, research…
TechImage-Bench: Rubric-Based Evaluation for Technical Image Generation
Minheng Ni, Zhengyuan Yang, Yaowen Zhang +9
We study technical image generation, where a model must synthesize information-dense, scientifically precise illustrations from detailed descriptions rather than merely produce vis…