6 papers
AgentBrew: Lifelong Knowledge Brewing from Strong Teachers to Weak LLM Agents
Yangqin Jiang, Chao Huang
Deploying LLM agents typically requires a compact test-time student, even if a stronger teacher is available during training. We study knowledge brewing: distilling a teacher's int…
AOHP: An Open-Source OS-Level Agent Harness for Personalized, Efficient and Secure Interaction
Shanhui Zhao, Jiacheng Liu, Guohong Liu +13
AI agents are driving a new software paradigm, with the ability to autonomously call tools, extract information, manage memory, and complete tasks that span applications and data s…
FastCode: Fast and Cost-Efficient Code Understanding and Reasoning
Zhonghang Li, Zongwei Li, Yuxuan Chen +5
Repository-scale code reasoning is a cornerstone of modern AI-assisted software engineering, enabling Large Language Models (LLMs) to handle complex workflows from program comprehe…
M3-AD: Reflection-aware Multi-modal, Multi-category, and Multi-dimensional Benchmark and Framework for Industrial Anomaly Detection
Chao Huang, Yanhui Li, Yunkang Cao +5
Although multimodal large language models (MLLMs) have advanced industrial anomaly detection toward a zero-shot paradigm, they still tend to produce high-confidence yet unreliable…
OpenPhone: Mobile Agentic Foundation Models
Yangqin Jiang, Chao Huang
With the advancement of multimodal large language models (MLLMs), building GUI agent systems has become an increasingly promising direction--especially for mobile platforms, given…
AI-Trader: Benchmarking Autonomous Agents in Real-Time Financial Markets
Tianyu Fan, Yuhao Yang, Yangqin Jiang +3
Large Language Models (LLMs) have demonstrated remarkable potential as autonomous agents, approaching human-expert performance through advanced reasoning and tool orchestration. Ho…