activity
20222026
most citedViWOZ: A Multi-Domain Task-Oriented Dialogue Systems Dataset For Low-resource Language

2 citations · 2 across the 4 of their papers we have counts for

collaborators

5 papers

cs.CV2026

Classroom Behavior Monitoring with YOLO An Empirical Study in Higher Education Settings

Sinh Vu Trong, Dung Nguyen Manh, Hieu Hoang Minh +3

Classroom behavior monitoring plays a vital role in evaluating student engagement and improving teaching effectiveness. Traditional observation methods remain subjective and lack s…

cs.SE2025

SWE-EVO: Benchmarking Coding Agents in Long-Horizon Software Evolution Scenarios

Tue Le, Minh V. T. Thai, Dung Nguyen Manh +2

Existing benchmarks for AI coding agents focus on isolated, single-issue tasks such as fixing a bug or adding a small feature. However, real-world software engineering is a long-ho…

cs.SE2024

CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs

Dung Nguyen Manh, Thang Phan Chau, Nam Le Hai +4

Recent advances in Code Large Language Models (CodeLLMs) have primarily focused on open-ended code generation, often overlooking the crucial aspect of code understanding and reason…

cs.CL2023

The Vault: A Comprehensive Multilingual Dataset for Advancing Code Understanding and Generation

Dung Nguyen Manh, Nam Le Hai, Anh T. V. Dau +4

We present The Vault, a dataset of high-quality code-text pairs in multiple programming languages for training large language models to understand and generate code. We present met…

cs.CL20222 cited

ViWOZ: A Multi-Domain Task-Oriented Dialogue Systems Dataset For Low-resource Language

Phi Nguyen Van, Tung Cao Hoang, Dung Nguyen Manh +2

Most of the current task-oriented dialogue systems (ToD), despite having interesting results, are designed for a handful of languages like Chinese and English. Therefore, their per…