5 papers
Soft Token Alignment for Cross-Lingual Reasoning
Jiayi He, Jungsoo Park, Wei Xu +1
Multilingual large language models often produce inconsistent reasoning and answers for semantically equivalent prompts in different languages. Prior work suggests that intermediat…
MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts
Jiayi He, Yangmin Huang, Qianyun Du +5
Deploying Large Language Models (LLMs) in medical applications requires fact-checking capabilities to ensure patient safety and regulatory compliance. We introduce MedFact, a chall…
Forward-Only Continual Learning
Jiao Chen, Jiayi He, Fangfang Chen +2
Catastrophic forgetting remains a central challenge in continual learning (CL) with pre-trained models. While existing approaches typically freeze the backbone and fine-tune a smal…
Self-Correction is More than Refinement: A Learning Framework for Visual and Language Reasoning Tasks
Jiayi He, Hehai Lin, Qingyun Wang +2
While Vision-Language Models (VLMs) have shown remarkable abilities in visual and language reasoning tasks, they invariably generate flawed responses. Self-correction that instruct…
Towards General Industrial Intelligence: A Survey of Continual Large Models in Industrial IoT
Jiao Chen, Jiayi He, Fangfang Chen +6
Industrial AI is transitioning from traditional deep learning models to large-scale transformer-based architectures, with the Industrial Internet of Things (IIoT) playing a pivotal…