6 papers
Large Language Models as Virtual Survey Respondents: Evaluating Sociodemographic Response Generation
Jianpeng Zhao, Chenyu Yuan, Weiming Luo +6
Questionnaire-based surveys are foundational to social science research and public policymaking, yet traditional survey methods remain costly, time-consuming, and often limited in…
STEMVerse: A Dual-Axis Diagnostic Framework for STEM Reasoning in Large Language Models
Xuzhao Li, Xuchen Li, Jian Zhao +1
As Large Language Models (LLMs) achieve significant breakthroughs in complex reasoning tasks, evaluating their proficiency in science, technology, engineering, and mathematics (STE…
Beyond Accuracy: Evaluating Grounded Visual Evidence in Thinking with Images
Xuchen Li, Xuzhao Li, Renjie Pi +3
Despite the remarkable progress of Vision-Language Models (VLMs) in adopting "Thinking-with-Images" capabilities, accurately evaluating the authenticity of their reasoning process…
Read the Docs Before Rewriting: Equip Rewriter with Domain Knowledge via Continual Pre-training
Qi Wang, Yixuan Cao, Yifan Liu +2
A Retrieval-Augmented Generation (RAG)-based question-answering (QA) system enhances a large language model's knowledge by retrieving relevant documents based on user queries. Disc…
Semantic Parsing for Question Answering over Knowledge Graphs
Sijia Wei, Wenwen Zhang, Qisong Li +1
In this paper, we propose a novel method for question answering over knowledge graphs based on graph-to-segment mapping, designed to improve the understanding of natural language q…
Self-Evolving GPT: A Lifelong Autonomous Experiential Learner
Jinglong Gao, Xiao Ding, Yiming Cui +4
To improve the performance of large language models (LLMs), researchers have explored providing LLMs with textual task-solving experience via prompts. However, they rely on manual…