3 papers
cs.CL2026
Beyond Token-Level Policy Gradients for Complex Reasoning with Large Language Models
Mufan Xu, Kehai Chen, Xuefeng Bai +4
Existing policy-gradient methods for auto-regressive language models typically select subsequent tokens one at a time as actions in the policy. While effective for many generation…
cs.CL2025
Memory-augmented Query Reconstruction for LLM-based Knowledge Graph Reasoning
Mufan Xu, Gewen Liang, Kehai Chen +5
Large language models (LLMs) have achieved remarkable performance on knowledge graph question answering (KGQA) tasks by planning and interacting with knowledge graphs. However, exi…
cs.CL2024
LLM-based Discriminative Reasoning for Knowledge Graph Question Answering
Mufan Xu, Kehai Chen, Xuefeng Bai +3
Large language models (LLMs) based on generative pre-trained Transformer have achieved remarkable performance on knowledge graph question-answering (KGQA) tasks. However, LLMs ofte…