2 papers
cs.CV2026
PLaMo 2.1-VL Technical Report
Tommi Kerola, Yuya Masuda, Takashi Masuko +5
We introduce PLaMo 2.1-VL, a lightweight Vision Language Model (VLM) for autonomous devices, available in 8B and 2B variants and designed for local and edge deployment with Japanes…
cs.LG2026
CAT: Circular-Convolutional Attention for Sub-Quadratic Transformers
Yoshihiro Yamada
Transformers have driven remarkable breakthroughs in natural language processing and computer vision, yet their standard attention mechanism still imposes O(N^2) complexity, hinder…