works on

From the 1 of 31 linked papers with an AI index.

activity
20242026
collaborators
Showing cs.LGShow all

5 papers · 1 filter

cs.LG2026

Beyond the Capability Boundary: Zeroth-Order Optimization for Self-Evolving LLM Agents

Bingzhen Liu, Xiaomeng Fan, Yuwei Wu +4

Self-evolving methods improve the capabilities of LLM agents by sampling trajectories from the underlying LLMs and learning from these trajectories. However, these methods struggle…

cs.LG2025

Efficient Multi-turn RL for GUI Agents via Decoupled Training and Adaptive Data Curation

Pengxiang Li, Zechen Hu, Zirui Shang +15

Vision-language model (VLM) based GUI agents show promise for automating complex desktop and mobile tasks, but face significant challenges in applying reinforcement learning (RL):…

cs.LG2025

Curvature Learning for Generalization of Hyperbolic Neural Networks

Xiaomeng Fan, Yuwei Wu, Zhi Gao +2

Hyperbolic neural networks (HNNs) have demonstrated notable efficacy in representing real-world data with hierarchical structures via exploiting the geometric properties of hyperbo…

cs.LG2025

Large-Scale Riemannian Meta-Optimization via Subspace Adaptation

Peilin Yu, Yuwei Wu, Zhi Gao +2

Riemannian meta-optimization provides a promising approach to solving non-linear constrained optimization problems, which trains neural networks as optimizers to perform optimizati…

cs.LG2024

Residual Hyperbolic Graph Convolution Networks

Yangkai Xue, Jindou Dai, Zhipeng Lu +2

Hyperbolic graph convolutional networks (HGCNs) have demonstrated representational capabilities of modeling hierarchical-structured graphs. However, as in general GCNs, over-smooth…