works on

From the 1 of 32 linked papers with an AI index.

activity
20242026
most citedShoppingBench: A Real-World Intent-Grounded Shopping Benchmark for LLM-based Agents

1 citations · 1 across the 7 of their papers we have counts for

collaborators

32 papers

cs.AI2026

Seeing the End at Step Zero: Accelerating Diffusion MLLMs via MLP Sparsity-Aware Truncation

Qicheng Zhao, Qi Sun, Zheyu Yan

The paper presents Seer, a training‑free approach that detects the true end of generated sequences in diffusion multimodal large language models by monitoring MLP activation sparsi…

cs.AI2026

ResilPhase: Plug-and-Play Phase Mapping and Noise-Resilient Macro-Trajectory Extrapolation for Diffusion Acceleration

Qicheng Zhao, Yu Li, Qi Sun +1

The adoption of powerful diffusion models is hindered by their significant inference latency. Recent ``cache-then-forecast'' schemes alleviate this issue by accelerating DiTs using…

cs.LG2026

Sakana Fugu Technical Report

Yujin Tang, Edoardo Cetin, Jinglue Xu +11

The capabilities of frontier Large Language Models (LLMs) continue to advance, with different providers increasingly specializing in distinct domains. This raises a natural next ob…

cs.CL2026

Quality Over Clicks: Iterative Reinforcement Learning for Early-Stage E-Commerce Query Suggestion

Qi Sun, Kejun Xiao, Huaipeng Zhao +2

Existing dialogue systems rely on query suggestion to enhance user engagement. Recent approaches mainly optimize generative models using click-through rate (CTR) models to align wi…

cs.CL20261 cited

ShoppingBench: A Real-World Intent-Grounded Shopping Benchmark for LLM-based Agents

Jiangyuan Wang, Kejun Xiao, Qi Sun +4

Existing benchmarks in e-commerce primarily focus on basic user intents, such as finding or purchasing products. However, real-world users often pursue more complex goals, such as…

cs.AI2026

From "Weak" Signals to Strong Models: Preference Delta Aggregation with LoRA Merging

Qi Sun, Siyue Zhang, Yulin Chen +3

Training strong large language models (LLMs) requires high-quality supervision, which is often scarce. Recent work shows that paired preference data from weak-weaker model pairs (e…