Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
OpenEval: Benchmarking Chinese LLMs across Capability, Alignment and Safety
Chuang Liu, Linhao Yu, Jiaxuan Li +11
The rapid development of Chinese large language models (LLMs) poses big challenges for efficient LLM evaluation. While current initiatives have introduced new benchmarks or evaluat…
cs.CL2024
Identifying Multiple Personalities in Large Language Models with External Evaluation
Xiaoyang Song, Yuta Adachi, Jessie Feng +6
As Large Language Models (LLMs) are integrated with human daily applications rapidly, many societal and ethical concerns are raised regarding the behavior of LLMs. One of the ways…