1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Jia Liu, Yue Wang, Zhiqi Lin +3
Large language model fine-tuning techniques typically depend on extensive labeled data, external guidance, and feedback, such as human alignment, scalar rewards, and demonstration.…