3 papers
cs.CL2025
The Hallucination Tax of Reinforcement Finetuning
Linxin Song, Taiwei Shi, Jieyu Zhao
Reinforcement finetuning (RFT) has become a standard approach for enhancing the reasoning capabilities of large language models (LLMs). However, its impact on model trustworthiness…
cs.CV2025
Attributed Synthetic Data Generation for Zero-shot Domain-specific Image Classification
Shijian Wang, Linxin Song, Ryotaro Shimizu +2
Zero-shot domain-specific image classification is challenging in classifying real images without ground-truth in-domain training examples. Recent research involved knowledge from t…
cs.CV2021
Product Re-identification System in Fully Automated Defect Detection
Chenggui Sun, Li Bin Song
In this work, we introduce a method and present an improved neural work to perform product re-identification, which is an essential core function of a fully automated product defec…