9 papers
Characterizing Treatment-Context Medication Evidence Across Clinic Notes and Structured EHR Medication History
Mingyang Jiang, Congning Ni, Weixin Liu +1
Clinic notes and structured electronic health record (EHR) medication history often contain different medication information. Same-visit disagreement between these sources may resu…
DRIFT: Direct-Recursive Intervention-Conditioned Forecasting of ICU Physiological Trajectories
Weixin Liu, Juming Xiong, Congning Ni +4
Many time-series forecasts depend not only on prior observations but also on actions specified during the forecast period. In intensive care units (ICUs), future vital signs and la…
Coverage-Controlled Preference Mining from Noisy Claim Verification for Evidence-Grounded Generation
Weixin Liu, Congning Ni, Qingyuan Song +4
Evidence-grounded generation produces summaries whose claims should be supported by supplied evidence, but claim-level verifiers provide noisy feedback and can reward models that s…
CoRA: Confidence-Rationale Alignment for Reliable Chain-of-Thought Reasoning
Juming Xiong, Weixin Liu, Kevin Guo +9
Chain-of-thought (CoT) reasoning can improve LLM performance, but high answer confidence may be misleading when the accompanying CoT rationale is plausible yet incomplete or poorly…
RadOT-Eval: Auditable Structured-Evidence Transport for Radiology Report Evaluation
Weixin Liu, Juming Xiong, Yang Li +5
Automatic evaluation is critical for high-stakes text generation, where errors often involve omitted findings, hallucinated content, polarity reversals, location changes, uncertain…
Vectors Are Not Neutral: Sensitive-Information Inference from Exported LLM Representations in Summarization
Weixin Liu, Bowen Qu, Juming Xiong +3
Large language model (LLM) summarization systems may pass compact vector representations of private inputs to downstream retrieval, monitoring, audit, or analytic workflows. Even w…