7 papers
DeepSeekMath Meets Order Book: Group-Aware Policy Optimization for High-Frequency Directional Trading
Sayak Charabarty, Souradip Pal
This paper studies reinforcement learning for high-frequency trading on limit order books by pairing an Order-Flow-based state model with policy-gradient methods. Instead of value-…
Is Sliding Window All You Need? An Open Framework for Long-Sequence Recommendation
Sayak Chakrabarty, Souradip Pal
Long interaction histories are central to modern recommender systems, yet training with long sequences is often dismissed as impractical under realistic memory and latency budgets.…
Identifying and Mitigating Gender Cues in Academic Recommendation Letters: An Interpretability Case Study
Charlotte S. Alexander, Shane Storks, Souradip Pal +4
Letters of recommendation (LoRs) can carry patterns of implicitly gendered language that can inadvertently influence downstream decisions, e.g. in hiring and admissions. In this wo…
PixRec: Leveraging Visual Context for Next-Item Prediction in Sequential Recommendation
Sayak Chakrabarty, Souradip Pal
Large Language Models (LLMs) have recently shown strong potential for usage in sequential recommendation tasks through text-only models, which combine advanced prompt design, contr…
Time-Constrained Recommendations: Reinforcement Learning Strategies for E-Commerce
Sayak Chakrabarty, Souradip Pal
Unlike traditional recommendation tasks, finite user time budgets introduce a critical resource constraint, requiring the recommender system to balance item relevance and evaluatio…
MM-PoE: Multiple Choice Reasoning via. Process of Elimination using Multi-Modal Models
Sayak Chakrabarty, Souradip Pal
This paper introduces Multiple Choice Reasoning via. Process of Elimination using Multi-Modal models, herein referred to as Multi-Modal Process of Elimination (MM-PoE). This novel…