From the 1 of 2 linked papers with an AI index.
2 papers
cs.AI2026
GrandCode: Achieving Grandmaster Level in Competitive Programming via Agentic Reinforcement Learning
DeepReinforce Team, Ornith Team, Xiaoya Li +4
The paper presents GrandCode, a multi‑agent reinforcement learning system that integrates hypothesis generation, solving, test creation, and summarization modules, and uses a new A…
cs.CL2026
Template-assisted Contrastive Learning of Task-oriented Dialogue Sentence Embeddings
Minsik Oh, Jiwei Li, Guoyin Wang
Learning high quality sentence embeddings from dialogues has drawn increasing attentions as it is essential to solve a variety of dialogue-oriented tasks with low annotation cost.…