Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
DUEL: Adversarial Self-Play for Multimodal Reasoning
Lin Qiu, Hanqing Zeng, Yao Liu +3
Reinforcement learning (RL) has emerged as an effective paradigm for improving the reasoning capability of vision-language models (VLMs). However, RL-based optimization typically d…
cs.CV2026
Efficient Long-Context Modeling in Diffusion Language Models via Block Approximate Sparse Attention
Wenhu Zhang, Yiming Wu, Huanyu Wang +6
Diffusion Language Models (DLMs) enable globally coherent, bidirectional, and controllable text generation, offering advantages over traditional autoregressive LLMs, while scaling…