24 citations · 94 across the 35 of their papers we have counts for
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Enhancing Action and Ingredient Modeling for Semantically Grounded Recipe Generation
Guoshan Liu, Bin Zhu, Yian Li +3
Recent advances in Multimodal Large Language Models (MLMMs) have enabled recipe generation from food images, yet outputs often contain semantically incorrect actions or ingredients…
cs.CL2025
Benchmarking Gaslighting Negation Attacks Against Multimodal Large Language Models
Bin Zhu, Yinxuan Gui, Huiyan Qi +3
Multimodal Large Language Models (MLLMs) have exhibited remarkable advancements in integrating different modalities, excelling in complex understanding and generation tasks. Despit…
cs.CL2020★ 20 cited
Multi-modal Cooking Workflow Construction for Food Recipes
Liangming Pan, Jingjing Chen, Jianlong Wu +5
Understanding food recipe requires anticipating the implicit causal effects of cooking actions, such that the recipe can be converted into a graph describing the temporal workflow…