4 papers
GSAR: Goal-State-Anchor Rewards for Mobile GUI Agents with Self-Evolving Data Synthesis
Long Zhang, Yuhan Chen, Chaoran Zhang +7
Vision-Language Models (VLMs) based GUI agents stand to benefit significantly from online reinforcement learning (RL). However, their training is bottlenecked by two fundamental is…
Backdoor Samples Detection Based on Perturbation Discrepancy Consistency in Pre-trained Language Models
Zuquan Peng, Jianming Fu, Lixin Zou +3
The use of unvetted third-party and internet data renders pre-trained models susceptible to backdoor attacks. Detecting backdoor samples is critical to prevent backdoor activation…
DANCE: Resource-Efficient Neural Architecture Search with Data-Aware and Continuous Adaptation
Maolin Wang, Tianshuo Wei, Sheng Zhang +6
Neural Architecture Search (NAS) has emerged as a powerful approach for automating neural network design. However, existing NAS methods face critical limitations in real-world depl…
Flow Matching based Sequential Recommender Model
Feng Liu, Lixin Zou, Xiangyu Zhao +5
Generative models, particularly diffusion model, have emerged as powerful tools for sequential recommendation. However, accurately modeling user preferences remains challenging due…