Showing cs.IRShow all
2 papers · 1 filter
cs.IR2026
RubricRAG: Towards Interpretable and Reliable LLM Evaluation via Domain Knowledge Retrieval for Rubric Generation
Kaustubh D. Dhole, Eugene Agichtein
Large language models (LLMs) are increasingly evaluated and sometimes trained using automated graders such as LLM-as-judges that output scalar scores or preferences. While convenie…
cs.IR2025
Generative Product Recommendations for Implicit Superlative Queries
Kaustubh D. Dhole, Nikhita Vedula, Saar Kuzi +3
In Recommender Systems, users often seek the best products through indirect, vague, or under-specified queries, such as "best shoes for trail running". Such queries, also referred…