2 papers
cs.CV2025
Bringing The Consistency Gap: Explicit Structured Memory for Interleaved Image-Text Generation
Zeteng Lin, Xingxing Li, Wen You +4
Existing Vision Language Models (VLMs) often struggle to preserve logic, entity identity, and artistic style during extended, interleaved image-text interactions. We identify this…
cs.CL2025
Emission-GPT: A domain-specific language model agent for knowledge retrieval, emission inventory and data analysis
Jiashu Ye, Tong Wu, Weiwen Chen +11
Improving air quality and addressing climate change relies on accurate understanding and analysis of air pollutant and greenhouse gas emissions. However, emission-related knowledge…