5 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.LG2023
Chain-Of-Thought Prompting Under Streaming Batch: A Case Study
Yuxin Tang
Recently, Large Language Models (LLMs) have demonstrated remarkable capabilities. Chain-of-Thought (CoT) has been proposed as a way of assisting LLMs in performing complex reasonin…
cs.CL2023★ 5 cited
Compress, Then Prompt: Improving Accuracy-Efficiency Trade-off of LLM Inference with Transferable Prompt
Zhaozhuo Xu, Zirui Liu, Beidi Chen +5
While the numerous parameters in Large Language Models (LLMs) contribute to their superior performance, this massive scale makes them inefficient and memory-hungry. Thus, they are…