2 citations · 2 across the 1 of their papers we have counts for
2 papers
cs.CL2025
ACADATA: Parallel Dataset of Academic Data for Machine Translation
Iñaki Lacunza, Javier Garcia Gilabert, Francesca De Luca Fornaciari +4
We present ACADATA, a high-quality parallel dataset for academic translation, that consists of two subsets: ACAD-TRAIN, which contains approximately 1.5 million author-generated pa…
cs.CL2025★ 2 cited
Salamandra Technical Report
Aitor Gonzalez-Agirre, Marc Pàmies, Joan Llop +21
This work introduces Salamandra, a suite of open-source decoder-only large language models available in three different sizes: 2, 7, and 40 billion parameters. The models were trai…