◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Itziar González-Dios

3 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2

Across the 2 of 3 papers where every author was matched, so the position is known.

fields
  • cs.CL3
ORCID 0000-0003-1048-5403

identity via Semantic Scholar / OpenAlex

most citedThe BigScience ROOTS Corpus: A 1.6TB Composite Multilingual Dataset

65 citations · 68 across the 3 of their papers we have counts for

collaborators

3 papers

cs.CL2023★ 3 cited

This is not a Dataset: A Large Negation Benchmark to Challenge Large Language Models

Iker García-Ferrero, Begoña Altuna, Javier Álvez +2

Although large language models (LLMs) have apparently acquired a certain level of grammatical knowledge and the ability to make generalizations, they fail to interpret negation, a…

cs.CL2023

Easy-to-Read in Germany: A Survey on its Current State and Available Resources

Margot Madina, Itziar Gonzalez-Dios, Melanie Siegel

Easy-to-Read Language (E2R) is a controlled language variant that makes any written text more accessible through the use of clear, direct and simple language. It is mainly aimed at…

cs.CL2023★ 65 cited

The BigScience ROOTS Corpus: A 1.6TB Composite Multilingual Dataset

Hugo Laurençon, Lucile Saulnier, Thomas Wang +51

As language models grow ever larger, the need for large-scale high-quality text datasets has never been more pressing, especially in multilingual settings. The BigScience workshop,…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.