10 citations · 64 across the 27 of their papers we have counts for
27 papers
Pessimistic Value Iteration for Multi-Task Data Sharing in Offline Reinforcement Learning
Chenjia Bai, Lingxiao Wang, Jianye Hao +4
Offline Reinforcement Learning (RL) has shown promising results in learning a task-specific policy from a fixed dataset. However, successful offline RL often relies heavily on the…
Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human Feedback
Yifu Yuan, Jianye Hao, Yi Ma +6
Reinforcement Learning with Human Feedback (RLHF) has received significant attention for performing tasks without the need for costly manual reward design by aligning human prefere…
Enhancing Robotic Manipulation with AI Feedback from Multimodal Large Language Models
Jinyi Liu, Yifu Yuan, Jianye Hao +4
Recently, there has been considerable attention towards leveraging large language models (LLMs) to enhance decision-making processes. However, aligning the natural language text in…
Machine Learning Insides OptVerse AI Solver: Design Principles and Applications
Xijun Li, Fangzhou Zhu, Hui-Ling Zhen +23
In an era of digital ubiquity, efficient resource management and decision-making are paramount across numerous industries. To this end, we present a comprehensive study on the inte…
Rethinking Decision Transformer via Hierarchical Reinforcement Learning
Yi Ma, Chenjun Xiao, Hebin Liang +1
Decision Transformer (DT) is an innovative algorithm leveraging recent advances of the transformer architecture in reinforcement learning (RL). However, a notable limitation of DT…
A Circuit Domain Generalization Framework for Efficient Logic Synthesis in Chip Design
Zhihai Wang, Lei Chen, Jie Wang +7
Logic Synthesis (LS) plays a vital role in chip design -- a cornerstone of the semiconductor industry. A key task in LS is to transform circuits -- modeled by directed acyclic grap…