3 papers
cs.LG2026
ReCo: Reweighting GRPO Against Distributional Concentration
Junoh Park, Junseo Hwang, Wonguk Cho +1
Group Relative Policy Optimization (GRPO) has become a standard reinforcement learning method for post-training language models. Recent work shows that GRPO can reduce the base mod…
cs.LG2025
PiCa: Parameter-Efficient Fine-Tuning with Column Space Projection
Junseo Hwang, Wonguk Cho, Taesup Kim
Fine-tuning large foundation models is essential for building expert models tailored to specialized tasks and domains, but fully updating billions of parameters is computationally…
cs.CV2024
Hollowed Net for On-Device Personalization of Text-to-Image Diffusion Models
Wonguk Cho, Seokeon Choi, Debasmit Das +4
Recent advancements in text-to-image diffusion models have enabled the personalization of these models to generate custom images from textual prompts. This paper presents an effici…