From the 1 of 9 linked papers with an AI index.
9 papers
PreDiff-LM: Pretrained Discrete Masked Diffusion Language Modeling with Hybrid Attention
Zhengtao Yao, Runhao Li, Xupeng Chen +12
The paper presents PreDiff-LM, a discrete masked diffusion language model that retains causal attention on the prompt while applying bidirectional attention within masked targets,…
Less Data, Better Alignment: Data-Centric Multi-Evaluator Agreement for Preference Optimization
Zhengtao Yao, Runhao Li, Xupeng Chen +12
Research on preference optimization often varies the training objective while holding the data fixed. We instead ask whether a small, high-confidence set of on-policy responses can…
CodeRescue: Budget-Calibrated Recovery Routing for Coding Agents
Qijia He, Jiayi Cheng, Chenqian Le +8
Coding agents increasingly operate in executable environments where a failed attempt produces actionable feedback rather than merely an incorrect answer. Existing cost-aware system…
VehAnchor: Metadata-Free Metric Scale Recovery from Vehicle Cues in Aerial Imagery
Yifei Chen, Chenqian Le, Jiayi Cheng +1
Autonomous aerial robots operating in GPS-denied or communication-degraded environments frequently lose access to camera metadata and telemetry, leaving onboard perception systems…
Iterative Multimodal Retrieval-Augmented Generation for Medical Question Answering
Xupeng Chen, Binbin Shi, Chenqian Le +5
Medical retrieval-augmented generation (RAG) systems typically operate on text chunks extracted from biomedical literature, discarding the rich visual content (tables, figures, str…
Why Does Grounding Hurt Medical VQA? Benchmarking, Diagnosis, and Fine-Tuning of Vision-Language Models
Xupeng Chen, Binbin Shi, Chenqian Le +5
Vision-language models (VLMs) are increasingly applied to medical visual question answering (Med-VQA), yet whether they can \emph{localize} the evidence behind their answers---a pr…