works on

From the 1 of 41 linked papers with an AI index.

activity
20242026
collaborators

41 papers

cs.AI2026

Consilience: Conformally Calibrated Communication Control for Hidden-Profile Multi-Agent Reasoning

Abhijith Babu, Ramneet Kaur, Vishal Pramanik +7

Multi-agent LLM systems can improve reasoning by pooling diverse perspectives, but their effectiveness depends on coordinating communication, particularly in hidden-profile setting…

cs.LG2026

Vector Symbolic Policy Gradient

Ryozo Masukawa, Sanggeon Yun, SungHeon Jeong +6

We answer this question with Vector-Symbolic Policy Gradient (VSPG), a discrete-action actor that represents each action by a unit-norm hypervector and scores it by similarity to t…

cs.CR2026

Trident : How to Break Deep Reinforcement Learning Cyber Defenses (Agentic)

Ryozo Masukawa, Ian Bryant, Armita Kazeminajafabadi +6

Autonomous cyber defense systems based on Deep Reinforcement Learning (DRL) have attracted significant research attention, yet remain evaluated almost exclusively against static, h…

cs.LG2026

Models Can Model, But Can't Bind: Structured Grounding in Text-to-Optimization

Zhiqi Gao, Albert Ge, Alexander Berenbeim +2

The paper investigates why text‑to‑optimization models struggle to correctly ground problem data, introduces a benchmark (Text2Opt‑Bench) to study this, and proposes a binding‑focu…

cs.AI2026

HiComm: Hierarchical Communication for Multi-agent Reinforcement Learning

Runze Zhao, Dongruo Zhou, Sumit Kumar Jha +2

Cooperative multi-agent reinforcement learning (MARL) often relies on communication to mitigate partial observability, yet most existing protocols treat messages as flat dense vect…

cs.AI2026

From Actions to Understanding: Conformal Interpretability of Temporal Concepts in LLM Agents

Trilok Padhi, Ramneet Kaur, Krishiv Agarwal +9

Large Language Models (LLMs) are increasingly deployed as autonomous agents capable of reasoning, planning, and acting within interactive environments. Despite their growing capabi…