2 papers
cs.CV2026
TwinICL: Diagnosing Multimodal In-Context Learning through Paired Counterfactuals
Zihan Xue, Po-Yi Lu, Serhii Honcharenko +5
In-context learning (ICL) enables models to infer tasks from demonstrations, but existing benchmarks generally lack matched text and image versions needed to compare ICL performanc…
cs.CV2026
Pop-Up Distractions Reveal Bag-of-Events Behavior in Video Large Language Models
Oscar Chew, Serhii Honcharenko, Qian-Hui Chen +4
A key capability for video understanding is reliably linking subjects to events across time, yet whether Video Large Language Models (VideoLLMs) actually achieve this remains uncle…