6 papers
SW--Bench: Benchmarking Autonomous Software Agent Generation for Agentic Web
Linyao Chen, Bo Huang, Qinlao Zhao +13
The Agentic Web is emerging as a paradigm in which autonomous software agents interact with online resources and with each other to accomplish user goals. However, the capacity of…
Towards Efficient and Evidence-grounded Mobility Prediction with LLM-Driven Agent
Linyao Chen, Qinlao Zhao, Zechen Li +7
Individual-level mobility prediction is central to urban simulation, transportation planning, and policy analysis. Supervised sequence models achieve strong accuracy but require ta…
EnvX: Agentize Everything with Agentic AI
Linyao Chen, Zimian Peng, Yingxuan Yang +4
The widespread availability of open-source repositories has led to a vast collection of reusable software components, yet their utilization remains manual, error-prone, and disconn…
SensorLLM: Aligning Large Language Models with Motion Sensors for Human Activity Recognition
Zechen Li, Shohreh Deldari, Linyao Chen +2
We introduce SensorLLM, a two-stage framework that enables Large Language Models (LLMs) to perform human activity recognition (HAR) from sensor time-series data. Despite their stro…
CRAB: Cross-environment Agent Benchmark for Multimodal Language Model Agents
Tianqi Xu, Linyao Chen, Dai-Jie Wu +13
The development of autonomous agents increasingly relies on Multimodal Language Models (MLMs) to perform tasks described in natural language with GUI environments, such as websites…
Finding the Sweet Spot: Preference Data Construction for Scaling Preference Optimization
Yao Xiao, Hai Ye, Linyao Chen +4
Iterative data generation and model retraining are widely used to align large language models (LLMs). It typically involves a policy model to generate on-policy responses and a rew…