WikiHint: A Human-Annotated Dataset for Hint Ranking and Generation
arXiv:2412.01626 · doi:10.1145/3726302.3730284
Abstract
The use of Large Language Models (LLMs) has increased significantly with users frequently asking questions to chatbots. In the time when information is readily accessible, it is crucial to stimulate and preserve human cognitive abilities and maintain strong reasoning skills. This paper addresses such challenges by promoting the use of hints as an alternative or a supplement to direct answers. We first introduce a manually constructed hint dataset, WikiHint, which is based on Wikipedia and includes 5,000 hints created for 1,000 questions. We then finetune open-source LLMs for hint generation in answer-aware and answer-agnostic contexts. We assess the effectiveness of the hints with human participants who answer questions with and without the aid of hints. Additionally, we introduce a lightweight evaluation method, HintRank, to evaluate and rank hints in both answer-aware and answer-agnostic settings. Our findings show that (a) the dataset helps generate more effective hints, (b) including answer information along with questions generally improves the quality of generated hints, and (c) encoder-based models perform better than decoder-based models in hint ranking.
Accepted at SIGIR 2025
References in corpus (9)
- Gemini: A Family of Highly Capable Multimodal Models
- Human-Centred Learning Analytics and AI in Education: a Systematic Literature Review
- A Survey of Automated Programming Hint Generation -- The HINTS Framework
- Large Language Models Meet NLP: A Survey
- Multi-hop Question Answering
- TriviaHG: A Dataset for Automatic Hint Generation from Factoid Questions
- Exploring Hint Generation Approaches in Open-Domain Question Answering
- HintEval: An Open-Source Python Toolkit for Hint Generation and Hint Evaluation
- Navigating the Landscape of Hint Generation Research: From the Past to the Future