Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
DiG-bench: Discovery in Games
Ruairidh M. Battleday, Kai Sandbrink, Jimi Cullen-Drohan +13
Discovery---formulating novel generalizations---is a central part of the scientific process. Despite its importance, there is a gap in the current AI benchmark landscape, with few…
cs.AI2025
Affordable AI Assistants with Knowledge Graph of Thoughts
Maciej Besta, Lorenzo Paleari, Jia Hao Andrea Jiang +15
Large Language Models (LLMs) are revolutionizing the development of AI assistants capable of performing diverse tasks across domains. However, current state-of-the-art LLM-driven a…
cs.AI2025
Reasoning Language Models: A Blueprint
Maciej Besta, Julia Barth, Eric Schreiber +16
Reasoning language models (RLMs), also known as Large Reasoning Models (LRMs), such as OpenAI's o1 and o3, DeepSeek-R1, and Alibaba's QwQ, have redefined AI's problem-solving capab…