2 papers
cs.LG2026
Safe Langevin Soft Actor Critic
Mahesh Keswani, Samyak Jain, Raunak P. Bhattacharyya
Balancing reward and safety in constrained reinforcement learning remains challenging due to poor generalization from sharp value minima and inadequate handling of heavy-tailed ris…
cs.CL2025
Integrating Domain Knowledge for Financial QA: A Multi-Retriever RAG Approach with LLMs
Yukun Zhang, Stefan Elbl Droguett, Samyak Jain
This research project addresses the errors of financial numerical reasoning Question Answering (QA) tasks due to the lack of domain knowledge in finance. Despite recent advances in…