activity
20192022
most citedGalactica: A Large Language Model for Science

266 citations · 288 across the 7 of their papers we have counts for

collaborators
Showing cs.CLShow all

13 papers · 1 filter

cs.CL2022266 cited

Galactica: A Large Language Model for Science

Ross Taylor, Marcin Kardas, Guillem Cucurull +6

Information overload is a major obstacle to scientific progress. The explosive growth in scientific literature and data has made it ever harder to discover useful insights in a lar…

cs.CL20222 cited

Which Discriminator for Cooperative Text Generation?

Antoine Chaffin, Thomas Scialom, Sylvain Lamprier +4

Language models generate texts by successively predicting probability distributions for next tokens given past ones. A growing field of interest tries to leverage external informat…

cs.CL20213 cited

BEAMetrics: A Benchmark for Language Generation Evaluation Evaluation

Thomas Scialom, Felix Hill

Natural language processing (NLP) systems are increasingly trained to generate open-ended text rather than classifying between responses. This makes research on evaluation metrics…

cs.CL20211 cited

To Beam Or Not To Beam: That is a Question of Cooperation for Language GANs

Thomas Scialom, Paul-Alexis Dray, Sylvain Lamprier +2

Due to the discrete nature of words, language GANs require to be optimized from rewards provided by discriminator networks, via reinforcement learning methods. This is a much harde…

cs.CL2020

Toward Stance-based Personas for Opinionated Dialogues

Thomas Scialom, Serra Sinem Tekiroglu, Jacopo Staiano +1

In the context of chit-chat dialogues it has been shown that endowing systems with a persona profile is important to produce more coherent and meaningful conversations. Still, the…

cs.CL2020

Synthetic Data Augmentation for Zero-Shot Cross-Lingual Question Answering

Arij Riabi, Thomas Scialom, Rachel Keraron +3

Coupled with the availability of large scale datasets, deep learning architectures have enabled rapid progress on the Question Answering task. However, most of those datasets are i…