works on

From the 1 of 12 linked papers with an AI index.

collaborators

12 papers

astro-ph.CO2026

Forecast for the detectability of patchy hydrogen reionization in WEAVE-QSO measurements of the Lyman- forest power spectrum at redshift

Ke Ma, James S. Bolton, Vid Iršič +9

We present the first detailed forecasts for the detectability of patchy hydrogen reionization in the one-dimensional Ly forest power spectrum to be measured by the WEAVE-QSO sur…

cs.AI2026

DeepTravel: An End-to-End Agentic Reinforcement Learning Framework for Autonomous Travel Planning Agents

Yansong Ning, Rui Liu, Jun Wang +6

The paper introduces DeepTravel, an end‑to‑end reinforcement‑learning framework that trains autonomous travel‑planning agents to plan itineraries, invoke external tools, and self‑c…

cs.LG2026

Rethinking Zero-Shot Time Series Classification: From Task-specific Classifiers to In-Context Inference

Juntao Fang, Shifeng Xie, Shengbin Nie +7

The zero-shot evaluation of time series foundation models (TSFMs) for classification typically uses a frozen encoder followed by a task-specific classifier. However, this practice…

cs.LG2026

Compress the Easy, Explore the Hard: Difficulty-Aware Entropy Regularization for Efficient LLM Reasoning

Qin-Wen Luo, Sheng Ren, Xiang Chen +4

Chain-of-Thought (CoT) has substantially empowered Large Language Models (LLMs) to tackle complex reasoning tasks, yet the verbose nature of explicit reasoning steps incurs prohibi…

cs.AI2026

Agent-Omit: Adaptive Context Omission for Efficient LLM Agents

Yansong Ning, Jun Fang, Naiqiang Tan +1

Managing agent context (e.g., thought and observation) during multi-turn agent-environment interactions is an emerging strategy to improve agent efficiency. However, existing studi…

cs.CL2026

WavBench: Benchmarking Reasoning, Colloquialism, and Paralinguistics for End-to-End Spoken Dialogue Models

Yangzhuo Li, Shengpeng Ji, Yifu Chen +6

With the rapid integration of advanced reasoning capabilities into spoken dialogue models, the field urgently demands benchmarks that transcend simple interactions to address real-…