From the 1 of 9 linked papers with an AI index.
3 papers · 1 filter
ARC-TGI: Human-Validated Task Generators with Reasoning Chain Templates for ARC-AGI
Jens Lehmann, Syeda Khushbakht, Nikoo Salehfard +4
The Abstraction and Reasoning Corpus (ARC-AGI) probes few-shot abstraction and rule induction on small visual grids, but progress is difficult to measure on static collections of h…
ReFactX: Scalable Reasoning with Reliable Facts via Constrained Generation
Riccardo Pozzi, Matteo Palmonari, Andrea Coletta +3
Knowledge gaps and hallucinations are persistent challenges for Large Language Models (LLMs), which generate unreliable responses when lacking the necessary information to fulfill…
Aligning Knowledge Graphs and Language Models for Factual Accuracy
Nur A Zarin Nishat, Andrea Coletta, Luigi Bellomarini +3
Large language models like GPT-4, Gemini, and Claude have transformed natural language processing (NLP) tasks such as question answering, dialogue generation, summarization, and so…