2 papers
cs.CL2025
One-Pass to Reason: Token Duplication and Block-Sparse Mask for Efficient Fine-Tuning on Multi-Turn Reasoning
Ritesh Goru, Shanay Mehta, Prateek Jain
Fine-tuning Large Language Models (LLMs) on multi-turn reasoning datasets requires N (number of turns) separate forward passes per conversation due to reasoning token visibility co…
cs.LG2025
Benchmarking Anomaly Detection Algorithms: Deep Learning and Beyond
Shanay Mehta, Shlok Mehendale, Nicole Fernandes +3
Detection of anomalous situations for complex mission-critical systems hold paramount importance when their service continuity needs to be ensured. A major challenge in detecting a…