3 papers
cs.LG2026
Safe Langevin Soft Actor Critic
Mahesh Keswani, Samyak Jain, Raunak P. Bhattacharyya
Balancing reward and safety in constrained reinforcement learning remains challenging due to poor generalization from sharp value minima and inadequate handling of heavy-tailed ris…
cs.CL2025
Integrating Domain Knowledge for Financial QA: A Multi-Retriever RAG Approach with LLMs
Yukun Zhang, Stefan Elbl Droguett, Samyak Jain
This research project addresses the errors of financial numerical reasoning Question Answering (QA) tasks due to the lack of domain knowledge in finance. Despite recent advances in…
cs.LG2025
Bonsai: Gradient-free Graph Condensation for Node Classification
Mridul Gupta, Samyak Jain, Vansh Ramani +2
Graph condensation has emerged as a promising avenue to enable scalable training of GNNs by compressing the training dataset while preserving essential graph characteristics. Our s…