4 papers
Temporal Entailment Pretraining for Clinical Language Models over EHR Data
Tatsunori Tanaka, Fi Zheng, Kai Sato +3
Clinical language models have achieved strong performance on downstream tasks by pretraining on domain specific corpora such as discharge summaries and medical notes. However, most…
System-2 Mathematical Reasoning via Enriched Instruction Tuning
Huanqia Cai, Yijun Yang, Zhifeng Li
Solving complex mathematical problems via system-2 reasoning is a natural human skill, yet it remains a significant challenge for current large language models (LLMs). We identify…
UFO: Unified Fact Obtaining for Commonsense Question Answering
Zhifeng Li, Yifan Fan, Bowei Zou +1
Leveraging external knowledge to enhance the reasoning ability is crucial for commonsense question answering. However, the existing knowledge bases heavily rely on manual annotatio…
KEPR: Knowledge Enhancement and Plausibility Ranking for Generative Commonsense Question Answering
Zhifeng Li, Bowei Zou, Yifan Fan +1
Generative commonsense question answering (GenCQA) is a task of automatically generating a list of answers given a question. The answer list is required to cover all reasonable ans…