activity
20242026
collaborators

8 papers

cs.LG2026

Large Models for Time Series and Spatio-Temporal Data: A Survey and Outlook

Ming Jin, Yaxuan Kong, Yuxuan Liang +13

Temporal data, including time series and spatio-temporal data, are pervasive in real-world applications. Generated in massive volumes by physical and virtual sensors, they record d…

cs.SE2026

Beyond Retrieval: A Multitask Benchmark and Model for Code Search

Siqiao Xue, Zihan Liao, Jin Qin +4

Code search has usually been evaluated as first-stage retrieval, even though production systems rely on broader pipelines with reranking and developer-style queries. Existing bench…

cs.CV2026

LookBench: A Live and Holistic Open Benchmark for Fashion Image Retrieval

Gensmo. ai, Chao Gao, Siqiao Xue +4

In this paper, we present LookBench (We use the term "look" to reflect retrieval that mirrors how people shop -- finding the exact item, a close substitute, or a visually consisten…

cs.LG2026

QuitoBench: A High-Quality Open Time Series Forecasting Benchmark

Siqiao Xue, Zhaoyang Zhu, Wei Zhang +7

Time series forecasting is critical across finance, healthcare, and cloud computing, yet progress is constrained by a fundamental bottleneck: the scarcity of large-scale, high-qual…

cs.CV2025

DORAEMON: A Unified Library for Visual Object Modeling and Representation Learning at Scale

Ke Du, Yimin Peng, Chao Gao +2

DORAEMON is an open-source PyTorch library that unifies visual object modeling and representation learning across diverse scales. A single YAML-driven workflow covers classificatio…

cs.CL2025

FAMMA: A Benchmark for Financial Domain Multilingual Multimodal Question Answering

Siqiao Xue, Xiaojing Li, Fan Zhou +3

In this paper, we introduce FAMMA, an open-source benchmark for \underline{f}in\underline{a}ncial \underline{m}ultilingual \underline{m}ultimodal question \underline{a}nswering (QA…