2 papers
cs.CV2026
Domain Generalization via Text-Anchored Information Bottleneck
Eunyi Lyou, Yunjeong Choi, Junho Lee +1
Visual recognition models often fail when deployed in new environments. Domain Generalization (DG) addresses this by learning representations that remain invariant to environment-s…
cs.LG2026
QUATRO: Query-Adaptive Trust Region Policy Optimization for LLM Fine-tuning
Doyeon Lee, Eunyi Lyou, Hyunsoo Cho +3
GRPO-style reinforcement learning (RL)-based LLM fine-tuning algorithms have recently gained popularity. Relying on heuristic trust-region approximations, however, they can lead to…