Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Causal-Counterfactual RAG: The Integration of Causal-Counterfactual Reasoning into RAG
Harshad Khadilkar, Abhay Gupta
Large language models (LLMs) have transformed natural language processing (NLP), enabling diverse applications by integrating large-scale pre-trained knowledge. However, their stat…
cs.CL2024
Leveraging Domain Knowledge for Efficient Reward Modelling in RLHF: A Case-Study in E-Commerce Opinion Summarization
Swaroop Nath, Tejpalsingh Siledar, Sankara Sri Raghava Ravindra Muddu +8
Reinforcement Learning from Human Feedback (RLHF) has become a dominating strategy in aligning Language Models (LMs) with human values/goals. The key to the strategy is learning a…