8 papers
Beyond Journals: Rethinking Research Evaluation in Hungarian Computer Science
János Tapolcai, Márk Jelasity, Lajos Rónyai +3
This study examines the role of top-tier conference publications in Hungarian computer science research. We show that the national scientometric practice, which is currently journa…
WMAttack: Automated Attack Search for Adversarial Evaluation of World-Model Agents
Zhixiang Guo, Siyuan Liang, Shi Fu +4
Despite the growing use of world models as decision-making agents, their adversarial robustness remains underexplored due to the lack of dedicated automated evaluation methods. A k…
When World Models Dream Wrong: Physical-Conditioned Adversarial Attacks against World Models
Zhixiang Guo, Siyuan Liang, Andras Balogh +4
Generative world models (WMs) are increasingly used to synthesize controllable, sensor-conditioned driving videos, yet their reliance on physical priors exposes novel attack surfac…
Verification of the Implicit World Model in a Generative Model via Adversarial Sequences
András Balogh, Márk Jelasity
Generative sequence models are typically trained on sample sequences from natural or formal languages. It is a crucial question whether -- or to what extent -- sample-based trainin…
Detecting Semantic Backdoors in a Mystery Shopping Scenario
Arpad Berta, Gabor Danner, Istvan Hegedus +1
Detecting semantic backdoors in classification models--where some classes can be activated by certain natural, but out-of-distribution inputs--is an important problem that has rece…
On the Brittleness of LLMs: A Journey around Set Membership
Lea Hergert, Gábor Berend, Mario Szegedy +2
Large language models (LLMs) achieve superhuman performance on complex reasoning tasks, yet often fail on much simpler problems, raising concerns about their reliability and interp…