2 papers
cs.CV2026
MovieRecapsQA: A Multimodal Open-Ended Video Question-Answering Benchmark
Shaden Shaar, Bradon Thymes, Sirawut Chaixanien +2
Understanding real-world videos such as movies requires integrating visual and dialogue cues. Yet existing VideoQA benchmarks struggle to capture this multimodal reasoning and, giv…
cs.CL2025
Are Triggers Needed for Document-Level Event Extraction?
Shaden Shaar, Wayne Chen, Maitreyi Chatterjee +3
Most existing work on event extraction has focused on sentence-level texts and presumes the identification of a trigger-span -- a word or phrase in the input that evokes the occurr…