2 papers
cs.SE2026
RefusalBench: Why Refusal Rate Misranks Frontier LLMs on Biological Research Prompts
Lukas Weidener, Marko BrkiÄ, Mihailo JovanoviÄ +2
Frontier large language models are increasingly deployed as orchestration backbones for biological research workflows, yet no shared evidence base exists for comparing their refusa…
cs.AI2026
From Agent-Only Social Networks to Autonomous Scientific Research: Lessons from OpenClaw and Moltbook, and the Architecture of ClawdLab and Beach.Science
Lukas Weidener, Marko BrkiÄ, Phillip Lee +3
In January 2026, the open-source agent framework OpenClaw and the agent-only social network Moltbook produced a large-scale dataset of autonomous AI-to-AI interaction, attracting s…