18 citations · 46 across the 15 of their papers we have counts for
4 papers · 1 filter
Large Language Models with Controllable Working Memory
Daliang Li, Ankit Singh Rawat, Manzil Zaheer +5
Large language models (LLMs) have led to a series of breakthroughs in natural language processing (NLP), owing to their excellent understanding and generation abilities. Remarkably…
Two-stage LLM Fine-tuning with Less Specialization and More Generalization
Yihan Wang, Si Si, Daliang Li +5
Pretrained large language models (LLMs) are general purpose problem solvers applicable to a diverse set of tasks with prompts. They can be further improved towards a specific task…
XR Hackathon Going Online: Lessons Learned from a Case Study with Goethe-Institut
Wiesław Kopeć, Kinga Skorupska, Anna Jaskulska +6
In this article we report a case study of a Language and Culture-oriented transdisciplinary XR hackathon organized with Goethe-Institut. The hackathon was hosted as an online event…
Robust Distillation for Worst-class Performance
Serena Wang, Harikrishna Narasimhan, Yichen Zhou +3
Knowledge distillation has proven to be an effective technique in improving the performance a student model using predictions from a teacher model. However, recent work has shown t…