From the 1 of 1 linked paper with an AI index.
1 paper
Dayong Liu, Chao Xu, Weihong Chen +5
The paper introduces CFG-Bench, a benchmark of videos and QA pairs to evaluate how well multimodal language models can generate fine-grained action instructions and higher-order re…