works on

From the 1 of 9 linked papers with an AI index.

collaborators

9 papers

cs.CV2026

Test-Time Registers as Global Priors for Tokenized Image Generation

Cheng-Yao Hong, Yifan Wang, Yuewei Lin +1

Attention-based models often develop attention sinks, where a small number of tokens repeatedly attract attention and accumulate unusually large activations. In vision transformers…

cs.CV2026

Together, Then Apart: Balancing Alignment and Distinctiveness for Multimodal Survival Analysis

Wenjing Liu, Qin Ren, Wen Zhang +2

The paper introduces TTA, a framework that first aligns shared patterns across histopathology images and genomic data and then preserves modality‑specific information to improve ca…

cs.CV2026

VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning

Lingxiao Li, Yifan Wang, Xinyan Gao +3

Chain-of-Thought (CoT) prompting has proven remarkably effective for eliciting complex reasoning in large language models (LLMs). Yet, its potential in multimodal large language mo…

cs.LG2026

FreeBridge: Variational Schrödinger Bridges for Cellular Transition Dynamics

Xurui Wang, Qin Ren, Jun Ma +2

High-content imaging assays quantify cellular responses to chemical and genetic perturbations, yet continuous trajectories of individual cells are unobservable because cells are ch…

cs.CV2026

Supervise Less, See More: Training-free Nuclear Instance Segmentation with Prototype-Guided Prompting

Wen Zhang, Qin Ren, Wenjing Liu +2

Accurate nuclear instance segmentation is a pivotal task in computational pathology, supporting data-driven clinical insights and facilitating downstream translational applications…

cs.CV2026

Ouroboros: Single-step Diffusion Models for Cycle-consistent Forward and Inverse Rendering

Shanlin Sun, Yifan Wang, Hanwen Zhang +5

While multi-step diffusion models have advanced both forward and inverse rendering, existing approaches often treat these problems independently, leading to cycle inconsistency and…