2 papers
cs.CL2026
AstroMind: A High-Fidelity Benchmark for Spacecraft Behavior Reasoning Based on Large Language Models
Hao Liu, Siyuan Yang, Qinglei Hu +1
Understanding why a spacecraft maneuvers -- rather than simply that it did -- is an increasingly important problem for space domain awareness as Earth orbits grow crowded and conte…
cs.AI2026
SkillGraph: Graph Foundation Priors for LLM Agent Tool Sequence Recommendation
Hao Liu, Dongyu Li
LLM agents must select tools from large API libraries and order them correctly. Existing methods use semantic similarity for both retrieval and ordering, but ordering depends on in…