2 citations · 2 across the 3 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
Meta-Policy Reflexion: Reusable Reflective Memory and Rule Admissibility for Resource-Efficient LLM Agent
Chunlong Wu, Ye Luo, Zhibo Qu +1
Large language model (LLM) agents achieve impressive single-task performance but commonly exhibit repeated failures, inefficient exploration, and limited cross-task adaptability. E…
cs.AI2024★ 2 cited
OpenAI-o1 AB Testing: Does the o1 model really do good reasoning in math problem solving?
Leo Li, Ye Luo, Tingyou Pan
The Orion-1 model by OpenAI is claimed to have more robust logical reasoning capabilities than previous large language models. However, some suggest the excellence might be partial…