Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Retrieval, Scoring, and Decoding Shape Performance and Stability in LLM-based Conversational Recommendation
Ante Kapetanovic, Tomislav Duricic, Andro Mercep +1
Large language models (LLMs) are increasingly used as rerankers in conversational recommender systems, yet measured gains depend strongly on the retrieval and inference protocol. O…
cs.CL2026
Anchoring Bias in LLM-as-a-Judge Systems: Prior Scores Compromise Evaluation Independence
Ante Kapetanovic, Kemal Altwlkany, Andro Mercep +2
Large language models (LLMs) increasingly assess generated content, giving rise to the LLM-as-a-Judge paradigm. These systems now score outputs, filter content, and gate iterative…