2 papers
cs.LG2026
TD-DPO: Difference-Aware Preference Optimization for Mitigating Sycophancy in Clinical Autism Intervention Dialogue
Shuzhong Lai, Junhong Lai, Chenxi Li +5
The sycophancy of large language models can increase the safety risk in intervention dialogue for autistic children. Supervised fine-tuning can somewhat reduce sycophancy, but rely…
cs.LG2025
Human-like Cognitive Generalization for Large Models via Brain-in-the-loop Supervision
Jiaxuan Chen, Yu Qi, Yueming Wang +1
Recent advancements in deep neural networks (DNNs), particularly large-scale language models, have demonstrated remarkable capabilities in image and natural language understanding.…