distributed execution 1GPU memory efficiency 1large language model training 1long-context reinforcement learning 1policy optimization 1
From the 1 of 7 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget
Changhai Zhou, Kieran Liu, Yuhua Zhou +17
LongStraw introduces an execution framework that enables reinforcement‑learning post‑training on million‑token prompts using a fixed GPU budget by separating prompt evaluation from…
cs.LG2025
ProtTeX-CC: Activating In-Context Learning in Protein LLM via Two-Stage Instruction Compression
Chuanliu Fan, Zicheng Ma, Jun Gao +5
Recent advances in protein large language models, such as ProtTeX, represent both side-chain amino acids and backbone structure as discrete token sequences of residue length. While…