Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Benchmarking Uncertainty Calibration in Large Language Model Long-Form Question Answering
Philip Müller, Nicholas PopoviÄ, Michael Färber +1
Large Language Models (LLMs) are commonly used in Question Answering (QA) settings, increasingly in the natural sciences if not science at large. Reliable Uncertainty Quantificatio…
cs.CL2024
Testing Uncertainty of Large Language Models for Physics Knowledge and Reasoning
Elizaveta Reganova, Peter Steinbach
Large Language Models (LLMs) have gained significant popularity in recent years for their ability to answer questions in various fields. However, these models have a tendency to "h…