collaborators

5 papers

cs.AI2026

ECPO: Evidence-Coupled Policy Optimization for Evidence-Certified Candidate Ranking

Miaobo Hu, Shuhao Hu, BoKun Wang +5

Ranking systems used in decision-support settings should not only order candidates but also expose evidence that can be independently checked. We study evidence-certified candidate…

cs.AI2026

SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text

Miaobo Hu, Xiaobo Guo, Shuhao Hu +5

Schema graphs are an upstream bottleneck of schema-grounded information extraction and knowledge graph construction, yet most extraction systems assume the schema is already availa…

cs.LG2026

AGPO: Adaptive Group Policy Optimization with Dual Statistical Feedback

Miaobo Hu, Shuhao Hu, Bokun Wang +5

Reinforcement learning improves LLM reasoning, but PPO/GRPO typically use fixed clipping and decoding temperature, which makes training brittle and tuning-heavy. We propose Adaptiv…

cs.CV2026

SAVER: Selective As-Needed Vision Evidence for Multimodal Information Extraction

Miaobo Hu, Shuhao Hu, Bokun Wang +5

Multimodal IE in social media is difficult because a post may attach multiple images that are weakly related, redundant, or even misleading with respect to the text. In this settin…

cs.LG2026

From Sparsity to Simplicity: Enabling Simpler Sequential Replacements via Sparse Attention Distillation

Yuxin Ren, Maxwell D Collins, Miao Hu +1

Self-attention serves as the core foundation of large-scale transformer pretraining, but its quadratic token interaction cost makes inference expensive. Replacing attention with si…