collaborators

9 papers

cs.LG2025

Divide-Fuse-Conquer: Eliciting "Aha Moments" in Multi-Scenario Games

Xiaoqing Zhang, Huabin Zheng, Ang Lv +5

Large language models (LLMs) have been observed to suddenly exhibit advanced reasoning abilities during reinforcement learning (RL), resembling an ``aha moment'' triggered by simpl…

cs.SI2025

The Stepwise Deception: Simulating the Evolution from True News to Fake News with LLM Agents

Yuhan Liu, Zirui Song, Juntian Zhang +3

With the growing spread of misinformation online, understanding how true news evolves into fake news has become crucial for early detection and prevention. However, previous resear…

cs.SE2025

Thinking Before Running! Efficient Code Generation with Thorough Exploration and Optimal Refinement

Xiaoqing Zhang, Yuhan Liu, Flood Sung +3

Code generation is crucial in software engineering for automating the coding process efficiently. While test-time computation methods show promise, they suffer from high latency du…

cs.LG2025

More is not always better? Enhancing Many-Shot In-Context Learning with Differentiated and Reweighting Objectives

Xiaoqing Zhang, Ang Lv, Yuhan Liu +6

Large language models (LLMs) excel at few-shot in-context learning (ICL) without requiring parameter updates. However, as ICL demonstrations increase from a few to many, performanc…

cs.SI2025

SAGraph: A Large-Scale Social Graph Dataset with Comprehensive Context for Influencer Selection in Marketing

Xiaoqing Zhang, Yuhan Liu, Jianzhou Wang +3

Influencer marketing campaign success heavily depends on identifying key opinion leaders who can effectively leverage their credibility and reach to promote products or services. T…

cs.RO2025

ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models

Zirui Song, Guangxian Ouyang, Mingzhe Li +10

Large Vision-Language Models (LVLMs) have recently advanced robotic manipulation by leveraging vision for scene perception and language for instruction following. However, existing…