works on

From the 1 of 10 linked papers with an AI index.

activity
20242026
collaborators

10 papers

cs.LG2026

The Seriality Gap in Video Diffusion Models

Jorge Diaz Chao, Konpat Preechakul, Yuxi Liu +1

The paper investigates why video diffusion models struggle with tasks that require sequential causal reasoning, such as multi‑ball collisions, and identifies a "seriality gap" wher…

cs.LG2026

The Serial Scaling Hypothesis

Yuxi Liu, Konpat Preechakul, Kananart Kuwaranancharoen +1

While machine learning has advanced through massive parallelization, we identify a critical blind spot: some problems are fundamentally sequential. These "inherently serial" proble…

cs.CV2026

ReBA-Pred-Net: Weakly-Supervised Regional Brain Age Prediction on MRI

Shuai Shao, Yan Wang, Shu Jiang +6

Brain age has become a prominent biomarker of brain health. Yet most prior work targets whole brain age (WBA), a coarse paradigm that struggles to support tasks such as disease cha…

cs.CV2025

Pillar-0: A New Frontier for Radiology Foundation Models

Kumar Krishna Agrawal, Longchao Liu, Long Lian +11

Radiology plays an integral role in modern medicine, yet rising imaging volumes have far outpaced workforce growth. Foundation models offer a path toward assisting with the full sp…

cs.CV2025

GRAID: Enhancing Spatial Reasoning of VLMs Through High-Fidelity Data Generation

Karim Elmaaroufi, Liheng Lai, Justin Svegliato +3

Vision Language Models (VLMs) achieve strong performance on many vision-language tasks but often struggle with spatial reasoning$\unicode{x2014}$a prerequisite for many application…

cs.LG2025

REOrdering Patches Improves Vision Models

Declan Kutscher, David M. Chan, Yutong Bai +2

Sequence models such as transformers require inputs to be represented as one-dimensional sequences. In vision, this typically involves flattening images using a fixed row-major (ra…