2 citations · 2 across the 4 of their papers we have counts for
5 papers
Classroom Behavior Monitoring with YOLO An Empirical Study in Higher Education Settings
Sinh Vu Trong, Dung Nguyen Manh, Hieu Hoang Minh +3
Classroom behavior monitoring plays a vital role in evaluating student engagement and improving teaching effectiveness. Traditional observation methods remain subjective and lack s…
SWE-EVO: Benchmarking Coding Agents in Long-Horizon Software Evolution Scenarios
Tue Le, Minh V. T. Thai, Dung Nguyen Manh +2
Existing benchmarks for AI coding agents focus on isolated, single-issue tasks such as fixing a bug or adding a small feature. However, real-world software engineering is a long-ho…
CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs
Dung Nguyen Manh, Thang Phan Chau, Nam Le Hai +4
Recent advances in Code Large Language Models (CodeLLMs) have primarily focused on open-ended code generation, often overlooking the crucial aspect of code understanding and reason…
The Vault: A Comprehensive Multilingual Dataset for Advancing Code Understanding and Generation
Dung Nguyen Manh, Nam Le Hai, Anh T. V. Dau +4
We present The Vault, a dataset of high-quality code-text pairs in multiple programming languages for training large language models to understand and generate code. We present met…
ViWOZ: A Multi-Domain Task-Oriented Dialogue Systems Dataset For Low-resource Language
Phi Nguyen Van, Tung Cao Hoang, Dung Nguyen Manh +2
Most of the current task-oriented dialogue systems (ToD), despite having interesting results, are designed for a handful of languages like Chinese and English. Therefore, their per…