collaborators

18 papers

cs.CL2026

DFM Mimir v1: An Open HRM Delivering Frontier Performance at 1B Parameters Using Only Permissible Post-Training Data

Peter Schneider-Kamp, Jacob Nielsen, Gianluca Barmina +2

Current large language model development relies on massive, often non-permissible datasets, creating a high barrier for researchers committed to open-source and ethically sourced d…

cs.MA2026

The Energy Society: A Simulation Environment for Studying Agent Cooperation under Survival Pressure

Lucas Bergholdt Hansen, Federico Torrielli, Filippo Tonini +1

The paper presents Energy Society, a minimal simulation where LLM-powered agents consume energy proportional to model size and must manage survival through jobs and donations, allo…

cs.IR2026

Influence of Prompt Engineering on Small Language Models for Guarded Query Routing

Richard Šléher, William Brach, Kristián Košťál +1

We study the problem of guarded query routing, where we assume that a user query first meets a router that either determines the ideal endpoint for in-distribution queries or rejec…

cs.CL2026

Training Language Models to Use Prolog as a Tool

Niklas Mellgren, Peter Schneider-Kamp, Lukas Galke Poech

Language models frequently produce plausible yet incorrect reasoning traces that are difficult to verify. We investigate fine-tuning models to use Prolog as an external symbolic re…

cs.AI2026

The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment

Filippo Tonini, Federico Torrielli, Anton Danholt Lautrup +3

As AI systems built from multiple language-model agents become more common, they are increasingly used to make decisions together: discussing, negotiating, and acting on shared tas…

cs.LG2026

BrainSurgery: Reproducible and Reliable Declarative Weight Manipulations for Model Editing and Upcycling

Gianluca Barmina, Annemette Broch Pirchert, Andrea Blasi Núñez +2

As deep learning models scale, managing, inspecting, and modifying large checkpoints has become increasingly challenging. Researchers often need to alter model weights for layer re…