3 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.HC2025
Interactive Reasoning: Visualizing and Controlling Chain-of-Thought Reasoning in Large Language Models
Rock Yuren Pang, K. J. Kevin Feng, Shangbin Feng +5
The output quality of large language models (LLMs) can be improved via "reasoning": generating segments of chain-of-thought (CoT) content to further condition the model prior to pr…
cs.CL2024
LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Weijia Shi, Xiaochuang Han, Chunting Zhou +4
We present LMFusion, a framework for empowering pretrained text-only large language models (LLMs) with multimodal generative capabilities, enabling them to understand and generate…
cs.CL2024★ 3 cited
MUSE: Machine Unlearning Six-Way Evaluation for Language Models
Weijia Shi, Jaechan Lee, Yangsibo Huang +7
Language models (LMs) are trained on vast amounts of text data, which may include private and copyrighted content. Data owners may request the removal of their data from a trained…