4 papers
Revisiting the Travel Planning Capabilities of Large Language Models
Bo-Wen Zhang, Jin Ye, Peng-Yu Hua +4
Travel planning serves as a critical task for long-horizon reasoning, exposing significant deficits in LLMs. However, existing benchmarks and evaluations primarily assess final pla…
A Systematic Comparison of Prompting and Multi-Agent Methods for LLM-based Stance Detection
Genan Dai, Zini Chen, Yi Yang +1
Stance detection identifies the attitude of a text author toward a given target. Recent studies have explored various LLM-based strategies for this task, from zero-shot prompting t…
ChinaTravel: An Open-Ended Travel Planning Benchmark with Compositional Constraint Validation for Language Agents
Jie-Jing Shao, Bo-Wen Zhang, Xiao-Wen Yang +8
Travel planning stands out among real-world applications of \emph{Language Agents} because it couples significant practical demand with a rigorous constraint-satisfaction challenge…
Neuro-Symbolic Artificial Intelligence: Towards Improving the Reasoning Abilities of Large Language Models
Xiao-Wen Yang, Jie-Jing Shao, Lan-Zhe Guo +5
Large Language Models (LLMs) have shown promising results across various tasks, yet their reasoning capabilities remain a fundamental challenge. Developing AI systems with strong r…