1 paper · 1 filter
Ahsan Bilal, Muhammad Ahmed Mohsin, Muhammad Umer +2
This survey explores the development of meta-thinking capabilities in Large Language Models (LLMs) from a Multi-Agent Reinforcement Learning (MARL) perspective. Meta-thinking self-…