From the 2 of 5 linked papers with an AI index.
5 papers
MultivationBench: A Benchmark for Multimodal Sequential Motivation Reasoning
Kawai Chung, Chunkit Chan, Yauwai Yim +12
The paper introduces MultivationBench, a benchmark that tests multimodal large language models on their ability to reason about evolving human motivations across sequential visual…
Explaining and Tuning Transformer-based LLMs in Arithmetic Tasks with Human Strategies
Luyu Qiu, Jianing Li, Hwanhee Kim +4
Transformer-based large language models (LLMs) continue to achieve state-of-the-art performance across various natural language processing tasks. However, their subpar performance…
Fine-grained CLIP fine-tuning with self-annotated region alignment
Chenyang Zhao, Wei Lin, Antoni B. Chan +1
The paper proposes SFF-CLIP, a fine-tuning approach that uses only image-text pairs to align region features with phrase concepts via text-specific heat maps, improving CLIP's fine…
Grad-ECLIP: Gradient-based Visual and Textual Explanations for CLIP
Chenyang Zhao, Kun Wang, Janet H. Hsiao +1
Significant progress has been achieved on the improvement and downstream usages of the Contrastive Language-Image Pre-training (CLIP) vision-language model, while less attention is…
Made-in China, Thinking in America:U.S. Values Persist in Chinese LLMs
David Haslett, Linus Ta-Lun Huang, Leila Khalatbari +2
As large language models increasingly mediate access to information and facilitate decision-making, they are becoming instruments in soft power competitions between global actors s…