6 papers
Balancing Understanding and Generation in Discrete Diffusion Models
Yue Liu, Yuzhong Zhao, Zheyong Xie +5
In discrete generative modeling, two dominant paradigms demonstrate divergent capabilities: Masked Diffusion Language Models (MDLM) excel at semantic understanding and zero-shot ge…
Vision Calorimeter for Anti-neutron Reconstruction: A Baseline
Hongtian Yu, Yangu Li, Mingrui Wu +8
In high-energy physics, anti-neutrons () are fundamental particles that frequently appear as final-state particles, and the reconstruction of their kinematic properties pr…
Geometric-Mean Policy Optimization
Yuzhong Zhao, Yue Liu, Junpeng Liu +9
Group Relative Policy Optimization (GRPO) has significantly enhanced the reasoning capability of large language models by optimizing the arithmetic mean of token-level rewards. Unf…
Building Vision Models upon Heat Conduction
Zhaozhi Wang, Yue Liu, Yunjie Tian +3
Visual representation models leveraging attention mechanisms are challenged by significant computational overhead, particularly when pursuing large receptive fields. In this study,…
CC-Diff: Enhancing Contextual Coherence in Remote Sensing Image Synthesis
Mu Zhang, Yunfan Liu, Yue Liu +2
Existing image synthesis methods for natural scenes focus primarily on foreground control, often reducing the background to simplistic textures. Consequently, these approaches tend…
VMamba: Visual State Space Model
Yue Liu, Yunjie Tian, Yuzhong Zhao +6
Designing computationally efficient network architectures remains an ongoing necessity in computer vision. In this paper, we adapt Mamba, a state-space language model, into VMamba,…