2 papers
cs.CL2025
Guiding Reasoning in Small Language Models with LLM Assistance
Yujin Kim, Euiin Yi, Minu Kim +2
The limited reasoning capabilities of small language models (SLMs) cast doubt on their suitability for tasks demanding deep, multi-step logical deduction. This paper introduces a f…
cs.LG2024
Preference Alignment with Flow Matching
Minu Kim, Yongsik Lee, Sehyeok Kang +3
We present Preference Flow Matching (PFM), a new framework for preference-based reinforcement learning (PbRL) that streamlines the integration of preferences into an arbitrary clas…