cs.AI
42.9k resultsEmergence Invariance: From Symbolized Thought to Interface Refinement
Yi Liu
Securing Agentic AI: From Per-Action Checks to Trajectory Assurance
Alireza Lotfi, Subangkar Karmaker Shanto, Imtiaz Karim +1
Does the Competitive Component of Adversarial Self-Play Improve Legal Reasoning? A Controlled Negative Result
Miseog Shawn Kim
Is More Privileged Information Better? From Solution Traces to Problem-Solving Structure in Self-Distilled Reasoning
Xuyang Zhao, Liting Zhang, Zichen Xu +4
Latent Thought Credit: Multi-Answer Credit Assignment for Latent Reasoning
Xuyang Zhao, Liting Zhang, Zichen Xu +4
Post-Training on Office Work Improves Software Engineering: A Behavioral Account of Cross-Domain Transfer
Logan Ritchie, Sushant Mehta, Liudas Panavas +1
When Memory Updates but Behavior Does Not: Repairing Implicit Stale Dependencies in Personalized Agent Responses
Haofei Sun, Lin He
Salami Attack: Stealthy Collusive Memory Poisoning against OpenClaw
Zheng Lin, Yuzhe Huang, Zhenxing Niu +2
GISAgentBench: A Practitioner-Sourced Benchmark for Evaluating LLM Agents on GIS Tasks
Abhinav Pothuri, Zhe Jiang, Zelin Xu +1
LongCat Sparse Attention: Taming the Lightning via Streaming-aware Hierarchical Cross-Layer Indexing
Wen Zan, Jiaqi Zhang, Jianchao Tan +11
Allocation Before Ranking: Decoupled Token Compression for OmniLLMs
Zhenghui Guo, Yilin Yang, Yuanbin Man +5
TCPO: Turn-Level Credit Policy Optimization
Sicong Liao, Zhi Chen, Yaohua Tang
When Memory Becomes Authority: Benchmarking Authority Collapse at the Memory Consolidation Boundary
Qiuyang Zhan, Rui Zhang, Sheng Guo +2
GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks
Jiarui Tan, Zhongjian Zhang, YaBo Guo +5
Beyond Single-Use Tokens: Durable Authorization State for Replay-Resistant LLM Agent Actions
Jinghan Xu, Longze Fan, Zeyuan Wang +2
Constructing Executable Analytical Knowledge Representations for Meta-Analysis Synthesis Using an Agentic Harness
Lingbo Li, Anuradha Mathrani, Teo Susnjak
LaCache: Robust Semantic Caching for LLM Serving
Jiacheng Liang, Yuhui Wang, Tanqiu Jiang +1
DAPD: Dual-Anchored Policy Distillation
Jianyu Wu, Yizhou Wang, Encheng Su +2
CoEvo-Mem: Co-Evolving Retrieval Policy and Memory Bank for LLM Agents
Bowen Ye, Yongchao Xu, Zhijian Li +3
MemSIF: From Structured Interactions to Dual-Track Fact Memory for LLM Agents
YuFei Luo, Xiucheng Xu, Zhen Yang
RL-Lock: Reinforcement Learning for Generating Interlocking Assemblies
Xuyang Ma, Chaewoon Kim, Haonan Zhang +3
Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs
Zixuan Huang, Yang Zhou, Kaixuan Wang +6
Leveraging AI for fine-grained food safety risk forecasting in sparse data conditions
Dongqi Wang, Weiwei Chen, Han Zhou +1
FRAMES: Guarded and Dual-Objective Skill Evolution for Agents in Policy-Governed Enterprise Workflows
Xuhui Wang, Ruoqi Shu, Chen Dan +4