3 papers
cs.CV2026
FinMTM: A Multi-Turn Multimodal Benchmark for Financial Reasoning and Agent Evaluation
Chenxi Zhang, Ziliang Gan, Liyun Zhu +3
The financial domain poses substantial challenges for vision-language models (VLMs) due to specialized chart formats and knowledge-intensive reasoning requirements. However, existi…
cs.CV2025
Retrospective Memory for Camouflaged Object Detection
Chenxi Zhang, Jiayun Wu, Qing Zhang +2
Camouflaged object detection (COD) primarily focuses on learning subtle yet discriminative representations from complex scenes. Existing methods predominantly follow the parametric…
cs.CV2025
CGCOD: Class-Guided Camouflaged Object Detection
Chenxi Zhang, Qing Zhang, Jiayun Wu +1
Camouflaged Object Detection (COD) aims to identify objects that blend seamlessly into their surroundings. The inherent visual complexity of camouflaged objects, including their lo…