activity
20242026
most citedBRIDGE: Benchmarking Large Language Models for Understanding Real-world Clinical Practice Text

3 citations · 3 across the 1 of their papers we have counts for

collaborators

5 papers

cs.CL2026

FinTruthQA: A Benchmark for AI-Driven Financial Disclosure Quality Assessment in Investor -- Firm Interactions

Peilin Zhou, Ziyue Xu, Xinyu Shi +7

Accurate and transparent financial information disclosure is essential for market efficiency, investor decision-making, and corporate governance. Chinese stock exchanges' investor…

cs.CL20263 cited

BRIDGE: Benchmarking Large Language Models for Understanding Real-world Clinical Practice Text

Jiageng Wu, Bowen Gu, Ren Zhou +14

Large language models (LLMs) hold great promise for medical applications and are evolving rapidly, with new models being released at an accelerated pace. However, benchmarking on l…

cs.CL2025

Why Chain of Thought Fails in Clinical Text Understanding

Jiageng Wu, Kevin Xie, Bowen Gu +3

Large language models (LLMs) are increasingly being applied to clinical care, a domain where both accuracy and transparent reasoning are critical for safe and trustworthy deploymen…

cs.CL2025

Scalable Medication Extraction and Discontinuation Identification from Electronic Health Records Using Large Language Models

Chong Shao, Douglas Snyder, Chiran Li +7

Identifying medication discontinuations in electronic health records (EHRs) is vital for patient safety but is often hindered by information being buried in unstructured notes. Thi…

cs.CL2024

Revealing COVID-19's Social Dynamics: Diachronic Semantic Analysis of Vaccine and Symptom Discourse on Twitter

Zeqiang Wang, Jiageng Wu, Yuqi Wang +5

Social media is recognized as an important source for deriving insights into public opinion dynamics and social impacts due to the vast textual data generated daily and the 'uncons…