Publications (22)
MARS: Unleashing the Power of Speculative Decoding via Margin-Aware Verification
Jingwei Song, Xinyu Wang, Hanbin Wang +6
Group Pattern Selection Optimization: Let LRMs Pick the Right Pattern for Reasoning
Hanbin Wang, Jingwei Song, Jinpeng Li +2
KnowCoder-X: Boosting Multilingual Information Extraction via Code
Yuxin Zuo, Wenxuan Jiang, Wenxuan Liu +7
COAST: Enhancing the Code Debugging Ability of LLMs through Communicative Agent Based Data Synthesis
Weiqing Yang, Hanbin Wang, Zhenghao Liu +7
INTERVENOR: Prompting the Coding Ability of Large Language Models with the Interactive Chain of Repair
Hanbin Wang, Zhenghao Liu, Shuo Wang +4
Observation of optical gyromagnetic properties in a magneto-plasmonic metamaterial
Weihao Yang, Qing Liu, Hanbin Wang +8
GameGPT: Multi-agent Collaborative Framework for Game Development
Dake Chen, Haoyang Zhang, Hanbin Wang +3
CODEMENV: Benchmarking Large Language Models on Code Migration
Keyuan Cheng, Xudong Shen, Yihao Yang +6
From and to : LLMs Learn New Skills in RL by Composing Old Ones
Lifan Yuan, Weize Chen, Yuchen Zhang +7
AADNet: Exploring EEG Spatiotemporal Information for Fast and Accurate Orientation and Timbre Detection of Auditory Attention Based on A Cue-Masked Paradigm
Keren Shi, Xu Liu, Xue Yuan +6
Advancing LLM Reasoning Generalists with Preference Trees
Lifan Yuan, Ganqu Cui, Hanbin Wang +12
Building A Coding Assistant via the Retrieval-Augmented Language Model
Xinze Li, Hanbin Wang, Zhenghao Liu +6
Teaching Large Reasoning Models Effective Reflection
Hanbin Wang, Jingwei Song, Jinpeng Li +5
A geometric reformulation and structure-preserving rotational discrete gradient scheme for the full Ericksen-Leslie model
Hanbin Wang, Jie Xu, Zhiguo Yang
Circular displacement current induced anomalous magneto-optical effects in high index Mie resonators
Shuang Xia, Daria Ignatyeva, Qing Liu +11
A second-order SO(3)-preserving and energy-stable scheme for orthonormal frame gradient flow model of biaxial nematic liquid crystals
Hanbin Wang, Jie Xu, Zhiguo Yang
Process Reinforcement through Implicit Rewards
Ganqu Cui, Lifan Yuan, Zefan Wang +22
Switching the Optical Chirality in Magneto-plasmonic Metasurfaces Using Applied Magnetic Fields
Jun Qin, Longjiang Deng, Tongtong Kang +11
UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning
Haoming Wang, Haoyang Zou, Huatong Song +109
Code-Vision: Evaluating Multimodal LLMs Logic Understanding and Code Generation Capabilities
Hanbin Wang, Xiaoxuan Zhou, Zhipeng Xu +7
In situ high resolution real-time quantum efficiency imaging for photocathodes
Dai Wu, Dexin Xiao, Jianxin Wang +16
Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence
Guanting Dong, Junting Lu, Junjie Huang +17