2 papers
cs.AI2026
What the Reranker Sees: Multi-Aspect Page Annotation for Long-Document Multimodal Question Answering
Guanchen Wu, Jiayuan Ding, Subhabrata Mukherjee +1
Long-document visual question answering (VQA) over documents of tens to hundreds of pages mixing text, tables, charts, and figures typically follows retrieve-then-read pipelines. I…
cs.LG2026
Predictive single cell foundation model for gene regulation and aging with privacy-preserving tabular learning
Jiayuan Ding, Jianhui Lin, Ziyang Miao +14
Pre-trained foundation models (FMs) have begun transforming single-cell genomics, but scaling them raises privacy concerns. Moreover, unlike text data, single-cell data is unordere…