From the 1 of 5 linked papers with an AI index.
5 papers
FOLIO: Focused Semantic Memory for Streaming Video Understanding
Haoyang Fan, Dhruv Parikh, Anvitha Ramachandran +4
The paper introduces FOLIO, a training‑free focused semantic memory system that records detailed information about important entities in a streaming video while compactly storing s…
Vision Non-Causal Trapezoidal Mamba: Eliminating Directional Scanning in Vision SSMs with Second-Order Dynamics
Anvitha Ramachandran, Dhruv Parikh, Haoyang Fan +2
State Space Models (SSMs) have emerged as an alternative to Vision Transformers, yet most vision SSMs inherit directional token scanning from causal sequence modeling. While effect…
Can Graphs Help Vision SSMs See Better?
Dhruv Parikh, Anvitha Ramachandran, Haoyang Fan +3
Vision state space models inherit the efficiency and long-range modeling ability of Mamba-style selective scans. However, their performance depends critically on the representation…
GraphLeap: Decoupling Graph Construction and Convolution for Vision GNN Acceleration on FPGA
Anvitha Ramachandran, Dhruv Parikh, Viktor Prasanna
Vision Graph Neural Networks (ViGs) represent an image as a graph of patch tokens, enabling adaptive, feature-driven neighborhoods. Unlike CNNs with fixed grid biases or Vision Tra…
Accelerating Dynamic Image Graph Construction on FPGA for Vision GNNs
Anvitha Ramachandran, Dhruv Parikh, Viktor Prasanna
Vision Graph Neural Networks (Vision GNNs, or ViGs) represent images as unstructured graphs, achieving state of the art performance in computer vision tasks such as image classific…