1 paper
Ingrid Petrova, Luan Vejsiu
Large autoregressive language models exhibit a self-correction blind spot: they reliably fix identical errors when attributed to an external source yet fail to fix the same errors…