4 papers
Trust the uncertain teacher: distilling dark knowledge via calibrated uncertainty
Jeonghyun Kim, SooKyung Kim, Richeng Xuan +1
The core of knowledge distillation lies in transferring the teacher's rich 'dark knowledge'-subtle probabilistic patterns that reveal how classes are related and the distribution o…
QUATRO: Query-Adaptive Trust Region Policy Optimization for LLM Fine-tuning
Doyeon Lee, Eunyi Lyou, Hyunsoo Cho +3
GRPO-style reinforcement learning (RL)-based LLM fine-tuning algorithms have recently gained popularity. Relying on heuristic trust-region approximations, however, they can lead to…
Post-Training and Test-Time Scaling of Generative Agent Behavior Models for Interactive Autonomous Driving
Hyunki Seong, Jeong-Kyun Lee, Heesoo Myeong +5
Learning interactive motion behaviors among multiple agents is a core challenge in autonomous driving. While imitation learning models generate realistic trajectories, they often i…
CountSteer: Steering Attention for Object Counting in Diffusion Models
Hyemin Boo, Hyoryung Kim, Myungjin Lee +4
Text-to-image diffusion models generate realistic and coherent images but often fail to follow numerical instructions in text, revealing a gap between language and visual represent…