NewEvery arXiv paper, its researchers & institutions — mapped.
papers

Publications (13)

cs.CV2026

ReMoT: Reinforcement Learning with Motion Contrast Triplets

Cong Wan, Zeyu Guo, Jiangyang Li +5

cs.AI2026

Kairos: A Regret-Aware Native World-Action Model Stack for Physical AI

Kairos Team, Fei Wang, Shan You +21

cs.IR2020

A novel sentence embedding based topic detection method for micro-blog

Cong Wan, Shan Jiang, Cuirong Wang +4

cs.CV2026

Trajectory-Diversity-Driven Robust Vision-and-Language Navigation

Jiangyang Li, Cong Wan, SongLin Dong +4

cs.RO2026

Retrieve-then-Steer: Online Success Memory for Test-Time Adaptation of Generative VLAs

Jianchao Zhao, Huoren Yang, Yusong Hu +6

cs.LG2026

DataClaw0: Agentic Tailoring Multimodal Data from Raw Streams

Cong Wan, Zeyu Guo, Zijian Cai +6

cs.CV2025

CanonSwap: High-Fidelity and Consistent Video Face Swapping via Canonical Space Modulation

Xiangyang Luo, Ye Zhu, Yunfei Liu +5

cs.CV2026

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning

Cong Wan, Ying He, Zhongzhan Huang +1

cs.CV2024

Prompt-Agnostic Adversarial Perturbation for Customized Diffusion Models

Cong Wan, Yuhang He, Xiang Song +1

cs.CV2026

ProSR: Process-Shaped Spatial Reasoning for Reliable Chain-of-Thought in VLMs

Jiangyang Li, Cong Wan, Changjie Wu +8

cs.CV2025

Grid: Omni Visual Generation

Cong Wan, Xiangyang Luo, Hao Luo +7

cs.CV2026

CoRe: A Comprehensive Framework for Cross-Image Comparative Reasoning in Vision-Language Models

Lin Peng, Cong Wan, Zeyu Guo +2

The paper introduces CoRe, a framework that improves vision-language models' ability to perform fine-grained cross‑image comparative reasoning by providing a large triplet‑based da…

#cross-image comparative reasoning#vision-language models#dataset construction#structured reward learning
physics.soc-ph2018

SI based disease model over signed network

Cong Wan, Cong Wang, Yanxia Lv