3 papers
cs.CV2026
Classroom Behavior Monitoring with YOLO An Empirical Study in Higher Education Settings
Sinh Vu Trong, Dung Nguyen Manh, Hieu Hoang Minh +3
Classroom behavior monitoring plays a vital role in evaluating student engagement and improving teaching effectiveness. Traditional observation methods remain subjective and lack s…
cs.SE2026
SWE-EVO: Benchmarking Coding Agents in Long-Horizon Software Evolution Scenarios
Tue Le, Minh V. T. Thai, Dung Nguyen Manh +2
Existing benchmarks for AI coding agents focus on isolated, single-issue tasks such as fixing a bug or adding a small feature. However, real-world software engineering is a long-ho…
cs.SE2025
CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs
Dung Nguyen Manh, Thang Phan Chau, Nam Le Hai +4
Recent advances in Code Large Language Models (CodeLLMs) have primarily focused on open-ended code generation, often overlooking the crucial aspect of code understanding and reason…