6 citations · 6 across the 1 of their papers we have counts for
4 papers
Characterizing Tradeoffs in Language Model Decoding with Informational Interpretations
Chung-Ching Chang, William W. Cohen, Yun-Hsuan Sung
We propose a theoretical framework for formulating language model decoder algorithms with dynamic programming and information theory. With dynamic programming, we lift the design o…
Memory Augmented Language Models through Mixture of Word Experts
Cicero Nogueira dos Santos, James Lee-Thorp, Isaac Noble +2
Scaling up the number of parameters of language models has proven to be an effective approach to improve performance. For dense models, increasing model size proportionally increas…
Hallucination Augmented Recitations for Language Models
Abdullatif Köksal, Renat Aksitov, Chung-Ching Chang
Attribution is a key concept in large language models (LLMs) as it enables control over information sources and enhances the factuality of LLMs. While existing approaches utilize o…
Characterizing Attribution and Fluency Tradeoffs for Retrieval-Augmented Large Language Models
Renat Aksitov, Chung-Ching Chang, David Reitter +2
Despite recent progress, it has been difficult to prevent semantic hallucinations in generative Large Language Models. One common solution to this is augmenting LLMs with a retriev…