2 citations · 3 across the 6 of their papers we have counts for
1 paper · 1 filter
Chenhong He, Lei Li, Shicheng Li +5
Hybrid attention dominates frontier LLMs, yet Vision Transformers (ViTs) in multimodal LLMs lack a satisfactory hybrid design, with no consensus on why certain attention patterns w…