1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2023
Visual Analytics for Efficient Image Exploration and User-Guided Image Captioning
Yiran Li, Junpeng Wang, Prince Aboagye +5
Recent advancements in pre-trained large-scale language-image models have ushered in a new era of visual comprehension, offering a significant leap forward. These breakthroughs hav…
cs.LG2023★ 1 cited
How Does Attention Work in Vision Transformers? A Visual Analytics Attempt
Yiran Li, Junpeng Wang, Xin Dai +5
Vision transformer (ViT) expands the success of transformer models from sequential data to images. The model decomposes an image into many smaller patches and arranges them into a…