3 papers
cs.CL2026
AgentLongBench: A Controllable Long Benchmark For Long-Contexts Agents via Environment Rollouts
Shicheng Fang, Yuxin Wang, Xiaoran Liu +6
The evolution of Large Language Models (LLMs) into autonomous agents necessitates the management of extensive, dynamic contexts. Current benchmarks, however, remain largely static,…
q-bio.BM2025
Few-shot Protein Fitness Prediction via In-context Learning and Test-time Training
Felix Teufel, Aaron W. Kollasch, Yining Huang +4
Accurately predicting protein fitness with minimal experimental data is a persistent challenge in protein engineering. We introduce PRIMO (PRotein In-context Mutation Oracle), a tr…
cs.LG2024
Multi-Scale Representation Learning for Protein Fitness Prediction
Zuobai Zhang, Pascal Notin, Yining Huang +5
Designing novel functional proteins crucially depends on accurately modeling their fitness landscape. Given the limited availability of functional annotations from wet-lab experime…