most citedImproving alignment of dialogue agents via targeted human judgements

133 citations · 138 across the 6 of their papers we have counts for

collaborators

6 papers

cs.LG20222 cited

Deep-Learning-Empowered Inverse Design for Freeform Reconfigurable Metasurfaces

Changhao Liu, Fan Yang, Maokun Li +1

The past decade has witnessed the advances of artificial intelligence with various applications in engineering. Recently, artificial neural network empowered inverse design for met…

stat.ML2022

SIMPLE-RC: Group Network Inference with Non-Sharp Nulls and Weak Signals

Jianqing Fan, Yingying Fan, Jinchi Lv +1

Large-scale network inference with uncertainty quantification has important applications in natural, social, and medical sciences. The recent work of Fan, Fan, Han and Lv (2022) in…

cs.CL2022

Structure-Unified M-Tree Coding Solver for MathWord Problem

Bin Wang, Jiangzhou Ju, Yang Fan +3

As one of the challenging NLP tasks, designing math word problem (MWP) solvers has attracted increasing research attention for the past few years. In previous work, models designed…

cs.LG2022133 cited

Improving alignment of dialogue agents via targeted human judgements

Amelia Glaese, Nat McAleese, Maja Trębacz +31

We present Sparrow, an information-seeking dialogue agent trained to be more helpful, correct, and harmless compared to prompted language model baselines. We use reinforcement lear…

math.LO2022

Intermediate logics in the setting of team semantics

Nick Bezhanishvili, Fan Yang

Several authors have recently defined intuitionistic logic based on team semantics (tIPC). In this paper we provide two alternative approaches to intermediate logics in the team se…

cs.LG20223 cited

ANT: Exploiting Adaptive Numerical Data Type for Low-bit Deep Neural Network Quantization

Cong Guo, Chen Zhang, Jingwen Leng +5

Quantization is a technique to reduce the computation and memory cost of DNN models, which are getting increasingly large. Existing quantization solutions use fixed-point integer o…