activity
20242026
collaborators

5 papers

cs.CL2026

Neural Grammatical Error Correction for Romanian

Teodor-Mihai Cotet, Stefan Ruseti, Mihai Dascalu

Resources for Grammatical Error Correction (GEC) in non-English languages are scarce, while available spellcheckers in these languages are mostly limited to simple corrections and…

cs.CL2026

Value-Aware Numerical Representations for Transformer Language Models

Andreea Dutulescu, Stefan Ruseti, Mihai Dascalu

Transformer-based language models often achieve strong results on mathematical reasoning benchmarks while remaining fragile on basic numerical understanding and arithmetic operatio…

cs.CL2026

Training Language Models with homotokens Leads to Delayed Overfitting

Adrian Cosma, Stefan Ruseti, Emilian Radoi +1

Subword tokenization introduces a computational layer in language models where many distinct token sequences decode to the same surface form and preserve meaning, yet induce differ…

cs.CL2025

The Strawberry Problem: Emergence of Character-level Understanding in Tokenized Language Models

Adrian Cosma, Stefan Ruseti, Emilian Radoi +1

Despite their remarkable progress across diverse domains, Large Language Models (LLMs) consistently fail at simple character-level tasks, such as counting letters in words, due to…

cs.CL2024

How Hard is this Test Set? NLI Characterization by Exploiting Training Dynamics

Adrian Cosma, Stefan Ruseti, Mihai Dascalu +1

Natural Language Inference (NLI) evaluation is crucial for assessing language understanding models; however, popular datasets suffer from systematic spurious correlations that arti…