#language models
38 papers · 1 filter
InfoOps Bench: A live information operations safety benchmark
Dorian Quelle, Lisa-Maria Neudert, Jonathan Bright +1
The paper introduces InfoOps Bench, a continuously updated benchmark that measures how easily frontier language models can be co-opted for state-backed information operations, usin…
Beyond Geometric Complementarity: Coherent Overlap in Sparse Mixture-of-Experts Routing
Huiyuan Tian, Bonan Xu, Shijian Li
The paper investigates how sparse mixture-of-experts language models route tokens to multiple experts, showing that expert subspaces overlap substantially yet routing still selects…
MORFES: A Benchmark for Productive Inflectional Competence in Modern Greek
Ioakeim Perros, Cleopatra Papadopoulou, Ayoub Kirouane +1
The paper presents MORFES, a benchmark of 500 expert‑verified items for testing Greek language models' ability to recognize and generate inflected word forms, especially for low‑fr…
Beyond Binary Rewards: A Comparative Study of Reward Design for Reinforcement Unlearning
Efstratios Zaradoukas, Davide Gabrielli, Bardh Prenkaj +1
The paper investigates how different reward functions affect the speed and effectiveness of reinforcement‑learning based machine unlearning for language models, proposing graded an…
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory
Rubin Wei, Jiaqi Cao, Jiarui Wang +4
The paper presents Memory Decoder at Scale, a pretrained parametric long‑term memory module for decoder‑only language models that is scaled up to 6.9 B parameters and shown to impr…
Hidden APIs in Language Models: Discovering Reusable Causal Interfaces from Forked Futures
SiYuan Ma, Yiqin Luo, Zhangji +8
The paper introduces a technique called forked futures that samples future operations after a prefix state to compare hidden states of language models, enabling the discovery of re…