Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Entropy-KL Divergence-based Token Masking: A Novel Approach for Selective Fine-tuning of Large Language Models
Qi Liu, Mingdi Sun, Yongyi He +5
Supervised fine-tuning (SFT) followed by reinforcement learning (RL) has become a standard post-training paradigm for large language models. This paradigm provides a cold-start for…
cs.AI2024
UniMEL: A Unified Framework for Multimodal Entity Linking with Large Language Models
Liu Qi, He Yongyi, Lian Defu +4
Multimodal Entity Linking (MEL) is a crucial task that aims at linking ambiguous mentions within multimodal contexts to the referent entities in a multimodal knowledge base, such a…
cs.AI2024
UniDM: A Unified Framework for Data Manipulation with Large Language Models
Yichen Qian, Yongyi He, Rong Zhu +8
Designing effective data manipulation methods is a long standing problem in data lakes. Traditional methods, which rely on rules or machine learning models, require extensive human…