1 citations · 1 across the 3 of their papers we have counts for
3 papers · 1 filter
ACR: Adaptive Context Refactoring via Context Refactoring Operators for Multi-Turn Dialogue
Jiawei Shen, Jia Zhu, Hanghui Guo +9
Large Language Models (LLMs) have shown remarkable performance in multi-turn dialogue. However, in multi-turn dialogue, models still struggle to stay aligned with what has been est…
Improving Autoregressive Training with Dynamic Oracles
Jianing Yang, Harshine Visvanathan, Yilin Wang +2
Many tasks within NLP can be framed as sequential decision problems, ranging from sequence tagging to text generation. However, for many tasks, the standard training methods, inclu…
Learning Mutually Informed Representations for Characters and Subwords
Yilin Wang, Xinyi Hu, Matthew R. Gormley
Most pretrained language models rely on subword tokenization, which processes text as a sequence of subword tokens. However, different granularities of text, such as characters, su…