Showing 2024 · cs.CLShow all
3 papers · 2 filters
cs.CL2024
2D-TPE: Two-Dimensional Positional Encoding Enhances Table Understanding for Large Language Models
Jia-Nan Li, Jian Guan, Wei Wu +2
Tables are ubiquitous across various domains for concisely representing structured information. Empowering large language models (LLMs) to reason over tabular data represents an ac…
cs.CL2024
Mixture-of-Modules: Reinventing Transformers as Dynamic Assemblies of Modules
Zhuocheng Gong, Ang Lv, Jian Guan +6
Is it always necessary to compute tokens from shallow to deep layers in Transformers? The continued success of vanilla Transformers and their variants suggests an undoubted "yes".…
cs.CL2024
From the Least to the Most: Building a Plug-and-Play Visual Reasoner via Data Synthesis
Chuanqi Cheng, Jian Guan, Wei Wu +1
We explore multi-step reasoning in vision-language models (VLMs). The problem is challenging, as reasoning data consisting of multiple steps of visual and language processing are b…