2 citations · 2 across the 1 of their papers we have counts for
1 paper
Sameera Horawalavithana, Sai Munikoti, Ian Stewart +2
Instruction finetuning is a popular paradigm to align large language models (LLM) with human intent. Despite its popularity, this idea is less explored in improving LLMs to align e…