Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Enhancing Action and Ingredient Modeling for Semantically Grounded Recipe Generation
Guoshan Liu, Bin Zhu, Yian Li +3
Recent advances in Multimodal Large Language Models (MLMMs) have enabled recipe generation from food images, yet outputs often contain semantically incorrect actions or ingredients…
cs.CL2025
Benchmarking Gaslighting Negation Attacks Against Multimodal Large Language Models
Bin Zhu, Yinxuan Gui, Huiyan Qi +3
Multimodal Large Language Models (MLLMs) have exhibited remarkable advancements in integrating different modalities, excelling in complex understanding and generation tasks. Despit…