◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Eunki Kim

8 papers hereh-index 440 citations10 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author6

Across the 8 of 8 papers where every author was matched, so the position is known.

fields
  • cs.CL3
  • cs.CV3
  • cs.AI1
  • cs.LG1

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.CLShow all

3 papers · 1 filter

cs.CL2025

Learning to Insert [PAUSE] Tokens for Better Reasoning

Eunki Kim, Sangryul Kim, James Thorne

To enhance reasoning capabilities, previous works have explored incorporating special-purpose tokens into the training process. These strategies strengthen the learning mechanism o…

cs.CL2025

AlphaPO: Reward Shape Matters for LLM Alignment

Aman Gupta, Shao Tang, Qingquan Song +10

Reinforcement Learning with Human Feedback (RLHF) and its variants have made huge strides toward the effective alignment of large language models (LLMs) to follow instructions and…

cs.CL2025

On the Robustness of Reward Models for Language Model Alignment

Jiwoo Hong, Noah Lee, Eunki Kim +5

The Bradley-Terry (BT) model is widely practiced in reward modeling for reinforcement learning with human feedback (RLHF). Despite its effectiveness, reward models (RMs) trained wi…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.