3 papers
cs.AI2026
Governed Shared Memory for Multi-Agent LLM Systems
Yanki Margalit, Nurit Cohen-Inger, Erni Avram +2
Multi-agent LLM environments require robust mechanisms for shared knowledge management. This paper formalizes the fleet-memory problem and identifies four foundational failure mode…
cs.AI2026
PeerRank: Autonomous LLM Evaluation Through Web-Grounded, Bias-Controlled Peer Review
Yanki Margalit, Erni Avram, Ran Taig +2
Evaluating large language models typically relies on human-authored benchmarks, reference answers, and human or single-model judgments, approaches that scale poorly, become quickly…
cs.SE2025
Black-Box Bug-Amplification for Multithreaded Software
Yeshayahu Weiss, Gal Amram, Achiya Elyasaf +3
Bugs, especially those in concurrent systems, are often hard to reproduce because they manifest only under rare conditions. Testers frequently encounter failures that occur only un…