3 citations · 3 across the 1 of their papers we have counts for
2 papers
cs.CL2025
Using Perspectival Words Is Harder Than Vocabulary Words for Humans and Even More So for Multimodal Language Models
Dota Tianai Dong, Yifan Luo, Po-Ya Angela Wang +2
Multimodal language models (MLMs) increasingly demonstrate human-like communication, yet their use of everyday perspectival words remains poorly understood. To address this gap, we…
cs.CV2023★ 3 cited
Vision-Language Integration in Multimodal Video Transformers (Partially) Aligns with the Brain
Dota Tianai Dong, Mariya Toneva
Integrating information from multiple modalities is arguably one of the essential prerequisites for grounding artificial intelligence systems with an understanding of the real worl…