1 citations · 2 across the 4 of their papers we have counts for
4 papers
Performance Prediction for Large Systems via Text-to-Text Regression
Yash Akhauri, Bryan Lewandowski, Cheng-Hsi Lin +7
In many industries, predicting metric outcomes of large systems is a fundamental problem, driven largely by traditional tabular regression. However, such methods struggle on comple…
Action-Dependent Optimality-Preserving Reward Shaping
Grant C. Forbes, Jianxun Wang, Leonardo Villalobos-Arias +2
Recent RL research has utilized reward shaping--particularly complex shaping rewards such as intrinsic motivation (IM)--to encourage agent exploration in sparse-reward environments…
Potential-Based Reward Shaping For Intrinsic Motivation
Grant C. Forbes, Nitish Gupta, Leonardo Villalobos-Arias +3
Recently there has been a proliferation of intrinsic motivation (IM) reward-shaping methods to learn in complex and sparse-reward environments. These methods can often inadvertentl…
Metric Ensembles For Hallucination Detection
Grant C. Forbes, Parth Katlana, Zeydy Ortiz
Abstractive text summarization has garnered increased interest as of late, in part due to the proliferation of large language models (LLMs). One of the most pressing problems relat…