3 papers
cs.AI2025
REAL: Benchmarking Abilities of Large Language Models for Housing Transactions and Services
Kexin Zhu, Yang Han
The development of large language models (LLMs) has greatly promoted the progress of chatbot in multiple fields. There is an urgent need to evaluate whether LLMs can play the role…
cs.CL2024
J2N -- Nominal Adjective Identification and its Application
Lemeng Qi, Yang Han, Zhuotong Xie
This paper explores the challenges posed by nominal adjectives (NAs) in natural language processing (NLP) tasks, particularly in part-of-speech (POS) tagging. We propose treating N…
cs.CL2024
Magnetic Preference Optimization: Achieving Last-iterate Convergence for Language Model Alignment
Mingzhi Wang, Chengdong Ma, Qizhi Chen +7
Self-play methods have demonstrated remarkable success in enhancing model capabilities across various domains. In the context of Reinforcement Learning from Human Feedback (RLHF),…