advantage shaping 1contrastive learning 1on-policy distillation 1policy optimization 1token-level correctness 1
From the 1 of 4 linked papers with an AI index.
Showing 2024Show all
2 papers · 1 filter
cs.CL2024
Natural Language Fine-Tuning
Jia Liu, Yue Wang, Zhiqi Lin +3
Large language model fine-tuning techniques typically depend on extensive labeled data, external guidance, and feedback, such as human alignment, scalar rewards, and demonstration.…
cs.HC2024
FaGeL: Fabric LLMs Agent empowered Embodied Intelligence Evolution with Autonomous Human-Machine Collaboration
Jia Liu, Min Chen
Recent advancements in Large Language Models (LLMs) have enhanced the reasoning capabilities of embodied agents, driving progress toward AGI-powered robotics. While LLMs have been…