4 papers
Training-Free Inference-Time Self-Reflection and Cost-Bounded Early Stopping for Large Language Models
Wei Yu, Suxing Liu, Minjie Yu +4
Reinforcement-learning training of reasoning LLMs (e.g., GRPO) is expensive and requires a controllable environment, committing every contribution to a full training pipeline. We p…
GRAFT: Biological Graph and Hypergraph Benchmarks for Linked Gene Expression and Phenotypic Trait Prediction in Arabidopsis thaliana
Manuel Serna-Aguilera, Vanshika Jindal, Fiona L. Goggin +5
Understanding which genes control which traits in an organism remains one of the central challenges in biology. Despite significant advances in data collection technology, our abil…
MetaResearcher: Scaling Deep Research via Self-Reflective Reinforcement Learning in Adversarial Virtual Environments
Wei Yu, Suxing Liu, Minjie Yu +4
Deep research agents have demonstrated remarkable capabilities in autonomous information gathering and synthesis, yet their training remains constrained by the static nature of sim…
AGP: A Novel Arabidopsis thaliana Genomics-Phenomics Dataset and its HyperGraph Baseline Benchmarking
Manuel Serna-Aguilera, Fiona L. Goggin, Aranyak Goswami +3
Understanding which genes control which traits in an organism remains one of the central challenges in biology. Despite significant advances in data collection technology, our abil…