2 papers
cs.CL2025
Token-level Accept or Reject: A Micro Alignment Approach for Large Language Models
Yang Zhang, Yu Yu, Bo Tang +8
With the rapid development of Large Language Models (LLMs), aligning these models with human preferences and values is critical to ensuring ethical and safe applications. However,…
cs.CL2025
Training Language Model to Critique for Better Refinement
Tianshu Yu, Chao Xiang, Mingchuan Yang +8
Large language models (LLMs) have demonstrated remarkable evaluation and critique capabilities, providing insightful feedback and identifying flaws in various tasks. However, limit…