3 papers
cs.LG2025
The Impact of Quantization on Large Reasoning Model Reinforcement Learning
Medha Kumar, Zifei Xu, Xin Wang +1
Strong reasoning capabilities can now be achieved by large-scale reinforcement learning (RL) without any supervised fine-tuning. Although post-training quantization (PTQ) and quant…
cs.DS2024
Optimized 2-Approximation of Treewidth
Mahdi Belbasi, Martin Fürer, Medha Kumar
This paper presents a linear FPT algorithm to find a tree decomposition with a 2-approximation of the treewidth with a significantly smaller exponential dependence on the treewidth…
cs.GT2024
The Degree of Fairness in Efficient House Allocation
Hadi Hosseini, Medha Kumar, Sanjukta Roy
The classic house allocation problem is primarily concerned with finding a matching between a set of agents and a set of houses that guarantees some notion of economic efficiency (…