23 papers
ASTELD: A Six-Axis Classification Framework for Autonomous AI Agents - Design, Evaluation, and an OpenClaw Case Study
Siyuan Li, Peng Shu, Churan Yu +19
Autonomous AI agent platforms differ substantially in architecture, security, tool integration, execution, autonomy, and deployment, yet the field lacks a common classification sch…
TraceLab: Characterizing Coding Agent Workloads for LLM Serving
Kan Zhu, Mathew Jacob, Chenxi Ma +4
Coding agents are rapidly becoming a major application of agentic LLMs, but serving them efficiently remains challenging. Progress on this challenge requires understanding real wor…
An Interaction Language Model: Mechanism Discovery from Statistical Patterns of Physical Interactions
Triparna Ganguly, Hanqi Jiang, Xinliang Li +5
Interactions among building blocks in physical, chemical, and biological systems follow structured patterns interpretable as a learnable language: just as language models learn whi…
MedVIGIL: Evaluating Trustworthy Medical VLMs Under Broken Visual Evidence
Hanqi Jiang, Junhao Chen, Mingyu Kang +12
Medical vision--language models (VLMs) are usually evaluated on intact image--question pairs, but trustworthy clinical use requires a stronger property: a model must recognise when…
Opportunities and Challenges of Large Language Models for Low-Resource Languages in Humanities Research
Tianyang Zhong, Zhenyuan Yang, Zhengliang Liu +11
Low-resource languages serve as invaluable repositories of human history, embodying cultural evolution and intellectual diversity. Despite their significance, these languages face…
Neural Functional Alignment Space: Brain-Referenced Representation of Artificial Neural Networks
Ruiyu Yan, Hanqi Jiang, Yi Pan +4
We propose the Neural Functional Alignment Space (NFAS), a brain-referenced representational framework for characterizing artificial neural networks on equal functional grounds. NF…