Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Reference-Free Rating of LLM Responses via Latent Information
Leander Girrbach, Chi-Ping Su, Tankred Saanum +3
How reliable are single-response LLM-as-a-judge ratings without references, and can we obtain fine-grained, deterministic scores in this setting? We study the common practice of as…
cs.CL2024
Dissecting Multiplication in Transformers: Insights into LLMs
Luyu Qiu, Jianing Li, Chi Su +2
Transformer-based large language models have achieved remarkable performance across various natural language processing tasks. However, they often struggle with seemingly easy task…