From the 1 of 9 linked papers with an AI index.
9 papers
Not All Retrievals are Useful: Cross-Attention for Input-Aware RAG in Time Series Forecasting
Seunghan Lee, Jaehoon Lee, Jun Seo +7
The paper introduces Cross-RAG, a retrieval-augmented generation framework for zero-shot time series forecasting that uses query‑retrieval cross‑attention to selectively attend to…
Pose-dIVE: Pose-Diversified Augmentation with Diffusion Model for Person Re-Identification
Inès Hyeonsu Kim, Woojeong Jin, Soowon Son +6
Person re-identification (Re-ID) often faces challenges due to variations in human poses and camera viewpoints, which significantly affect the appearance of individuals across imag…
Grounding World Simulation Models in a Real-World Metropolis
Junyoung Seo, Hyunwook Choi, Minkyung Kwon +10
What if a world simulation model could render not an imagined environment but a city that actually exists? Prior generative world models synthesize visually plausible yet artificia…
MoECLIP: Patch-Specialized Experts for Zero-shot Anomaly Detection
Jun Yeong Park, JunYoung Seo, Minji Kang +1
The CLIP model's outstanding generalization has driven recent success in Zero-Shot Anomaly Detection (ZSAD) for detecting anomalies in unseen categories. The core challenge in ZSAD…
3D Scene Prompting for Scene-Consistent Camera-Controllable Video Generation
JoungBin Lee, Jaewoo Jung, Jisang Han +6
We present 3DScenePrompt, a framework that generates the next video chunk from arbitrary-length input while enabling precise camera control and preserving scene consistency. Unlike…
CAMEO: Correspondence-Attention Alignment for Multi-View Diffusion Models
Minkyung Kwon, Jinhyeok Choi, Jiho Park +6
Multi-view diffusion models have recently emerged as a powerful paradigm for novel view synthesis, yet the underlying mechanism that enables their view-consistency remains unclear.…