activity
20172023
most citedBeyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

565 citations · 1.4k across the 18 of their papers we have counts for

collaborators
Showing 2023 · cs.CLShow all

6 papers · 2 filters

cs.CL2023

Evaluation of Faithfulness Using the Longest Supported Subsequence

Anirudh Mittal, Timo Schick, Mikel Artetxe +1

As increasingly sophisticated language models emerge, their trustworthiness becomes a pivotal issue, especially in tasks such as summarization and question-answering. Ensuring thei…

cs.CL2023★ 13 cited

Self-Alignment with Instruction Backtranslation

Xian Li, Ping Yu, Chunting Zhou +5

We present a scalable method to build a high quality instruction following language model by automatically labelling human-written text with corresponding instructions. Our approac…

cs.CL2023★ 1 cited

Active Learning Principles for In-Context Learning with Large Language Models

Katerina Margatina, Timo Schick, Nikolaos Aletras +1

The remarkable advancements in large language models (LLMs) have significantly enhanced the performance in few-shot learning settings. By using only a small number of labeled examp…

cs.CL2023★ 11 cited

LongForm: Effective Instruction Tuning with Reverse Instructions

Abdullatif Köksal, Timo Schick, Anna Korhonen +1

Instruction tuning enables language models to more effectively generalize and better follow user intent. However, obtaining instruction data is costly and challenging. Prior work e…

cs.CL2023★ 143 cited

Augmented Language Models: a Survey

Grégoire Mialon, Roberto Dessì, Maria Lomeli +10

This survey reviews works in which language models (LMs) are augmented with reasoning skills and the ability to use tools. The former is defined as decomposing a potentially comple…

cs.CL2023★ 400 cited

Toolformer: Language Models Can Teach Themselves to Use Tools

Timo Schick, Jane Dwivedi-Yu, Roberto Dessì +5

Language models (LMs) exhibit remarkable abilities to solve new tasks from just a few examples or textual instructions, especially at scale. They also, paradoxically, struggle with…