benchmarking 1conversation agents 1foundation models 1gradient-free updates 1LLM evaluation 1long-term memory 1memory-augmented models 1memory operations 1multimodal learning 1native memory 1
From the 2 of 15 linked papers with an AI index.
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
FTII-Bench: A Comprehensive Multimodal Benchmark for Flow Text with Image Insertion
Jiacheng Ruan, Yebin Yang, Zehao Lin +4
Benefiting from the revolutionary advances in large language models (LLMs) and foundational vision models, large vision-language models (LVLMs) have also made significant progress.…
cs.CV2024
MM-CamObj: A Comprehensive Multimodal Dataset for Camouflaged Object Scenarios
Jiacheng Ruan, Wenzhen Yuan, Zehao Lin +5
Large visual-language models (LVLMs) have achieved great success in multiple applications. However, they still encounter challenges in complex scenes, especially those involving ca…