activity
20232026
most citedLLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

38 citations · 67 across the 44 of their papers we have counts for

collaborators
Showing 2025Show all

9 papers · 1 filter

cs.AI2025

JustEva: A Toolkit to Evaluate LLM Fairness in Legal Knowledge Inference

Zongyue Xue, Siyuan Zheng, Shaochun Wang +8

The integration of Large Language Models (LLMs) into legal practice raises pressing concerns about judicial fairness, particularly due to the nature of their "black-box" processes.…

cs.CY2025★ 2 cited

Simulating Dispute Mediation with LLM-Based Agents for Legal Research

Junjie Chen, Haitao Li, Minghao Qin +6

Legal dispute mediation plays a crucial role in resolving civil disputes, yet its empirical study is limited by privacy constraints and complex multivariate interactions. To addres…

cs.CY2025

Chinese Court Simulation with LLM-Based Agent System

Kaiyuan Zhang, Jiaqi Li, Yueyue Wu +7

Mock trial has long served as an important platform for legal professional training and education. It not only helps students learn about realistic trial procedures, but also provi…

cs.CL2025

LLMs on Trial: Evaluating Judicial Fairness for Large Language Models

Yiran Hu, Zongyue Xue, Haitao Li +10

Large Language Models (LLMs) are increasingly used in high-stakes fields where their decisions impact rights and equity. However, LLMs' judicial fairness and implications for socia…

cs.IR2025

SelfRACG: Enabling LLMs to Self-Express and Retrieve for Code Generation

Qian Dong, Jia Chen, Qingyao Ai +6

Existing retrieval-augmented code generation (RACG) methods typically use an external retrieval module to fetch semantically similar code snippets used for generating subsequent fr…

cs.CL2025

Overview of the NTCIR-18 Automatic Evaluation of LLMs (AEOLLM) Task

Junjie Chen, Haitao Li, Zhumin Chu +2

In this paper, we provide an overview of the NTCIR-18 Automatic Evaluation of LLMs (AEOLLM) task. As large language models (LLMs) grow popular in both academia and industry, how to…