1 paper · 1 filter
Dingyi Kang, Dongming Jiang, Yi Li +2
Interaction between users and LLM agents is increasingly multimodal: conversations interleave text with images, and a later question may target either. Yet most agent memories are…