1 citations · 1 across the 1 of their papers we have counts for
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2024
When Every Token Counts: Optimal Segmentation for Low-Resource Language Models
Bharath Raj, Garvit Suri, Vikrant Dewangan +1
Traditional greedy tokenization methods have been a critical step in Natural Language Processing (NLP), influencing how text is converted into tokens and directly impacting model p…
cs.CL2024
SwitchCIT: Switching for Continual Instruction Tuning
Xinbo Wu, Max Hartman, Vidhata Arjun Jayaraman +1
Large language models (LLMs) and multimodal models (MMs) have exhibited impressive capabilities in various domains, particularly in general language understanding and visual reason…
cs.CL2024★ 1 cited
Transformer-based Causal Language Models Perform Clustering
Xinbo Wu, Lav R. Varshney
Even though large language models (LLMs) have demonstrated remarkable capability in solving various natural language tasks, the capability of an LLM to follow human instructions is…