2 papers
cs.AI2025
TGPR: Tree-Guided Policy Refinement for Robust Self-Debugging of LLMs
Daria Ozerova, Ekaterina Trofimova
Iterative refinement has been a promising paradigm to enable large language models (LLMs) to resolve difficult reasoning and problem-solving tasks. One of the key challenges, howev…
cs.CL2025
ML2B: Benchmarking LLMs on Cross-Lingual ML Pipeline Generation
Ekaterina Trofimova, Zosia Shamina, Maria Selifanova +7
We introduce ML2B, the first benchmark for evaluating cross-lingual task comprehension in end-to-end ML pipeline generation by large language models. Despite growing global AI adop…