10 papers
Self-Supervised Skill Optimization
Siran Peng, Cuiyu Yang, Tianyu Fu +9
Agent skills provide frozen large language model (LLM) agents with reusable procedural guidance, and recent work shows that such skills can be optimized with ground-truth (GT) feed…
Private Face Recognition Training Dataset Publication via Identity-Decoupled and Geometry-Preserving Face Distillation
Shuhuan Chen, Xiangyu Zhu, Weisong Zhao +6
The paper proposes a method to publish face recognition training datasets that removes personal identity information while keeping the geometric relationships needed for effective…
WebRetriever: A Large-Scale Comprehensive Benchmark for Efficient Web Agent Evaluation
Wei Dong, Tianyu Fu, Zhe Yu +9
As web agents increasingly demonstrate capabilities in automated task execution, the development of robust evaluation frameworks for assessing their navigation and task completion…
CPG-PAD: Concept-Informed Prompts Guided Presentation Attack Detection
Haoyuan Zhang, Xiangyu Zhu, Li Gao +3
Presentation Attack Detection (PAD) serves as a crucial safeguard for face recognition systems against presentation attacks such as printed photos, replayed videos, and 3D masks. D…
UPA: Unsupervised Prompt Agent via Tree-Based Search and Selection
Siran Peng, Weisong Zhao, Tianyu Fu +6
Prompt agents have recently emerged as a promising paradigm for automated prompt optimization, framing prompt discovery as a sequential decision-making problem over a structured pr…
One Ring to Rule Them All: Unifying Group-Based RL via Dynamic Power-Mean Geometry
Weisong Zhao, Tong Wang, Zichang Tan +11
Group-based reinforcement learning has evolved from the arithmetic mean of GRPO to the geometric mean of GMPO. While GMPO improves stability by constraining a conservative objectiv…