most citedOn the development of an AI performance and behavioural measures for teaching and classroom management

2 citations · 2 across the 6 of their papers we have counts for

collaborators

6 papers

cs.AI2026

SkillGLoW: Procedural-Family Skill Consolidation for Self-Improving Agents on Long-Horizon Task Streams

Ao Yan, Xin Zhang, Jiawei Du +1

LLM agents increasingly self-improve by writing and reusing textual skills, kept either as one global document or as a flat pool of per-task entries, though most of the evidence co…

cs.CV2026

Draw2Think: Harnessing Geometry Reasoning through Constraint Engine Interaction

Juncheng Hu, Jiawei Du, Xin Zhang +1

Vision-language models solve geometry problems with rising accuracy, yet their intermediate states remain latent and unverifiable: a relation expressed in textual reasoning or draw…

cs.LG2026

Visual Latents Know More Than They Say: Unsilencing Latent Reasoning in MLLMs

Xin Zhang, Qiqi Tao, Jiawei Du +2

Continuous latent-space reasoning offers a compact alternative to textual chain-of-thought for multimodal models, enabling high-dimensional visual evidence to be integrated without…

cs.AI2025

SAG-Agent: Enabling Long-Horizon Reasoning in Strategy Games via Dynamic Knowledge Graphs

Chenwei Tang, Lin Long, Xinyu Liu +6

Most commodity software lacks accessible Application Programming Interfaces (APIs), requiring autonomous agents to interact solely through pixel-based Graphical User Interfaces (GU…

cs.CV2025★ 2 cited

On the development of an AI performance and behavioural measures for teaching and classroom management

Andreea I. Niculescu, Jochen Ehnes, Chen Yi +9

This paper presents a two-year research project focused on developing AI-driven measures to analyze classroom dynamics, with particular emphasis on teacher actions captured through…

cs.AI2025

Rethinking Agent Design: From Top-Down Workflows to Bottom-Up Skill Evolution

Jiawei Du, Jinlong Wu, Yuzheng Chen +3

Most LLM-based agent frameworks adopt a top-down philosophy: humans decompose tasks, define workflows, and assign agents to execute each step. While effective on benchmark-style ta…