2 papers
cs.CV2025
Video Event Reasoning and Prediction by Fusing World Knowledge from LLMs with Vision Foundation Models
L'ea Dubois, Klaus Schmidt, Chengyu Wang +3
Current video understanding models excel at recognizing "what" is happening but fall short in high-level cognitive tasks like causal reasoning and future prediction, a limitation r…
cs.CL2024
Manual Verbalizer Enrichment for Few-Shot Text Classification
Quang Anh Nguyen, Nadi Tomeh, Mustapha Lebbah +3
With the continuous development of pre-trained language models, prompt-based training becomes a well-adopted paradigm that drastically improves the exploitation of models for many…