48 citations · 48 across the 2 of their papers we have counts for
2 papers
cs.CL2023★ 48 cited
InstructUIE: Multi-task Instruction Tuning for Unified Information Extraction
Xiao Wang, Weikang Zhou, Can Zu +11
Large language models have unlocked strong multi-task capabilities from reading instructive prompts. However, recent studies have shown that existing large models still have diffic…
stat.ML2022
Choquet regularization for reinforcement learning
Xia Han, Ruodu Wang, Xun Yu Zhou
We propose \emph{Choquet regularizers} to measure and manage the level of exploration for reinforcement learning (RL), and reformulate the continuous-time entropy-regularized RL pr…