activity
20242026
most citedJustice or Prejudice? Quantifying Biases in LLM-as-a-Judge

8 citations · 10 across the 7 of their papers we have counts for

collaborators

17 papers

cs.CL2026

CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era

Kaiwen Shi, Weixiang Sun, Zheyuan Zhang +3

Scientific research relies on citation integrity, yet large language models (LLMs) have introduced a critical risk: fabricated references that appear plausible but correspond to no…

cs.HC2025

From Verification Burden to Trusted Collaboration: Design Goals for LLM-Assisted Literature Reviews

Brenda Nogueira, Werner Geyer, Andrew Anderson +4

Large Language Models (LLMs) are increasingly embedded in academic writing practices. Although numerous studies have explored how researchers employ these tools for scientific writ…

cs.AI2025

The Reasoning Boundary Paradox: How Reinforcement Learning Constrains Language Models

Phuc Minh Nguyen, Chinh D. La, Duy M. H. Nguyen +3

Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a key method for improving Large Language Models' reasoning capabilities, yet recent evidence suggests it may p…

cs.CL2025

ChemOrch: Empowering LLMs with Chemical Intelligence via Synthetic Instructions

Yue Huang, Zhengzhe Jiang, Xiaonan Luo +12

Empowering large language models (LLMs) with chemical intelligence remains a challenge due to the scarcity of high-quality, domain-specific instruction-response datasets and the mi…

cs.CL2025

LLMs4All: A Review of Large Language Models Across Academic Disciplines

Yanfang Ye, Zheyuan Zhang, Tianyi Ma +26

Cutting-edge Artificial Intelligence (AI) techniques keep reshaping our view of the world. For example, Large Language Models (LLMs) based applications such as ChatGPT have shown t…

cs.CL2025

Breaking Language Barriers: Equitable Performance in Multilingual Language Models

Tanay Nagar, Grigorii Khvatskii, Anna Sokol +1

Cutting-edge LLMs have emerged as powerful tools for multilingual communication and understanding. However, LLMs perform worse in Common Sense Reasoning (CSR) tasks when prompted i…