3 citations · 3 across the 3 of their papers we have counts for
1 paper · 1 filter
Laurène Vaugrante, Francesca Carlon, Maluna Menke +1
Recent research on large language models (LLMs) has demonstrated their ability to understand and employ deceptive behavior, even without explicit prompting. However, such behavior…