1 paper · 1 filter
Runyan Tan, Shuang Wu, Phillip Howard
Obtaining high-quality outputs from Large Language Models (LLMs) often depends upon the choice of a sampling-based decoding strategy to probabilistically choose the next token at e…