activity
20192026
most citedSuccessive Over Relaxation Q-Learning

17 citations · 35 across the 8 of their papers we have counts for

collaborators

11 papers

cs.LG2026

Trust Guided Decision Transformer

Chainesh Gautam, Raghuram Bharadwaj Diddigi, Chandramouli Kamanchi +3

Decision Transformer performance degrades on long rollouts because the conditioning context drifts out of the training distribution. We show that this drift is visible through the…

cs.CV2024

Image Generation from Image Captioning -- Invertible Approach

Nandakishore S Menon, Chandramouli Kamanchi, Raghuram Bharadwaj Diddigi

Our work aims to build a model that performs dual tasks of image captioning and image generation while being trained on only one task. The central idea is to train an invertible mo…

cs.LG2024

Towards Unbiased Evaluation of Time-series Anomaly Detector

Debarpan Bhattacharya, Sumanta Mukherjee, Chandramouli Kamanchi +3

Time series anomaly detection (TSAD) is an evolving area of research motivated by its critical applications, such as detecting seismic activity, sensor failures in industrial plant…

cs.LG2024

Activations Through Extensions: A Framework To Boost Performance Of Neural Networks

Chandramouli Kamanchi, Sumanta Mukherjee, Kameshwaran Sampath +4

Activation functions are non-linearities in neural networks that allow them to learn complex mapping between inputs and outputs. Typical choices for activation functions are ReLU,…

stat.AP2020★ 1 cited

An Application of Newsboy Problem in Supply Chain Optimisation of Online Fashion E-Commerce

Chandramouli Kamanchi, Gopinath Ashok Kumar, Nachiappan Sundaram +2

We describe a supply chain optimization model deployed in an online fashion e-commerce company in India called Myntra. Our model is simple, elegant and easy to put into service. Th…

cs.LG2019★ 2 cited

A Convergent Off-Policy Temporal Difference Algorithm

Raghuram Bharadwaj Diddigi, Chandramouli Kamanchi, Shalabh Bhatnagar

Learning the value function of a given policy (target policy) from the data samples obtained from a different policy (behavior policy) is an important problem in Reinforcement Lear…