most citedClassMind: Scaling Classroom Observation and Instructional Feedback with Multimodal AI

1 citations · 1 across the 3 of their papers we have counts for

collaborators

6 papers

cs.CV2026

Envisioning global urban development with satellite imagery and generative AI

Kailai Sun, Yuebing Liang, Mingyi He +5

Urban development has been a defining force in human history, shaping cities for centuries. However, past studies mostly analyze such development as predictive tasks, failing to re…

cs.CV2026

Perception-Aware Multimodal Spatial Reasoning from Monocular Images

Yanchun Cheng, Rundong Wang, Xulei Yang +4

Spatial reasoning from monocular images is essential for autonomous driving, yet current Vision-Language Models (VLMs) still struggle with fine-grained geometric perception, partic…

cs.LG2026

Adaptive Problem Generation via Symbolic Representations

Teresa Yeo, Myeongho Jeon, Dulaj Weerakoon +4

We present a method for generating training data for reinforcement learning with verifiable rewards to improve small open-weights language models on mathematical tasks. Existing da…

cs.HC20251 cited

ClassMind: Scaling Classroom Observation and Instructional Feedback with Multimodal AI

Ao Qu, Yuxi Wen, Jiayi Zhang +6

Classroom observation -- one of the most effective methods for teacher development -- remains limited due to high costs and a shortage of expert coaches. We present ClassMind, an A…

cs.CL2025

MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents

Zijian Zhou, Ao Qu, Zhaoxuan Wu +6

Modern language agents must operate over long-horizon, multi-turn interactions, where they retrieve external information, adapt to observations, and answer interdependent queries.…

cs.CL2025

TETRIS: Optimal Draft Token Selection for Batch Speculative Decoding

Zhaoxuan Wu, Zijian Zhou, Arun Verma +3

We propose TETRIS, a novel method that optimizes the total throughput of batch speculative decoding in multi-request settings. Unlike existing methods that optimize for a single re…