Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
GuirlVG: Incentivize GUI Visual Grounding via Empirical Exploration on Reinforcement Learning
Weitai Kang, Bin Lei, Gaowen Liu +2
Graphical user interface visual grounding (GUI-VG), a core capability for GUI agents, has primarily relied on supervised fine-tuning (SFT) of multimodal large language models (MLLM…
cs.AI2024
Long text outline generation: Chinese text outline based on unsupervised framework and large language mode
Yan Yan, Yuanchi Ma
Outline generation aims to reveal the internal structure of a document by identifying underlying chapter relationships and generating corresponding chapter summaries. Although exis…