2 papers
cs.LG2026
Training Large Language Models for Self-Explanation Faithfulness
Yeoktatt Cheah, MarÃa Pérez-Ortiz, Noah Y. Siegel +1
We propose a Reinforcement Learning (RL) method to directly optimize the faithfulness of self-explanations - the extent to which a model's generated reasoning accurately reflects i…
cs.RO2025
OPA-Pack: Object-Property-Aware Robotic Bin Packing
Jia-Hui Pan, Yeok Tatt Cheah, Zhengzhe Liu +5
Robotic bin packing aids in a wide range of real-world scenarios such as e-commerce and warehouses. Yet, existing works focus mainly on considering the shape of objects to optimize…