2 papers
cs.AI2025
InMind: Evaluating LLMs in Capturing and Applying Individual Human Reasoning Styles
Zizhen Li, Chuanhao Li, Yibin Wang +8
LLMs have shown strong performance on human-centric reasoning tasks. While previous evaluations have explored whether LLMs can infer intentions or detect deception, they often over…
cs.AI2025
AI Idea Bench 2025: AI Research Idea Generation Benchmark
Yansheng Qiu, Haoquan Zhang, Zhaopan Xu +4
Large-scale Language Models (LLMs) have revolutionized human-AI interaction and achieved significant success in the generation of novel ideas. However, current assessments of idea…