activity
20242026
collaborators
Showing cs.LGShow all

5 papers · 1 filter

cs.LG2025

ACE-RLHF: Automated Code Evaluation and Socratic Feedback Generation Tool using Large Language Models and Reinforcement Learning with Human Feedback

Tasnia Rahman, Sathish A. P. Kumar, Sumit Jha +1

Automated Program Repair tools are developed for generating feedback and suggesting a repair method for erroneous code. State of the art (SOTA) code repair methods rely on data-dri…

cs.LG2025

DML-RAM: Deep Multimodal Learning Framework for Robotic Arm Manipulation using Pre-trained Models

Sathish Kumar, Swaroop Damodaran, Naveen Kumar Kuruba +2

This paper presents a novel deep learning framework for robotic arm manipulation that integrates multimodal inputs using a late-fusion strategy. Unlike traditional end-to-end or re…

cs.LG2025

Enhanced Penalty-based Bidirectional Reinforcement Learning Algorithms

Sai Gana Sandeep Pula, Sathish A. P. Kumar, Sumit Jha +1

This research focuses on enhancing reinforcement learning (RL) algorithms by integrating penalty functions to guide agents in avoiding unwanted actions while optimizing rewards. Th…

cs.LG2025

MORAL: A Multimodal Reinforcement Learning Framework for Decision Making in Autonomous Laboratories

Natalie Tirabassi, Sathish A. P. Kumar, Sumit Jha +1

We propose MORAL (a multimodal reinforcement learning framework for decision making in autonomous laboratories) that enhances sequential decision-making in autonomous robotic labor…

cs.LG2024

Automaton Distillation: Neuro-Symbolic Transfer Learning for Deep Reinforcement Learning

Suraj Singireddy, Precious Nwaorgu, Andre Beckus +5

Reinforcement learning (RL) is a powerful tool for finding optimal policies in sequential decision processes. However, deep RL methods have two weaknesses: collecting the amount of…