activity
20242026
collaborators

7 papers

cs.AI2026

The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling

Tu Nguyen, Matthieu Zimmer, Rasul Tutunov +2

A recurring pattern in "reasoning without training" is that base LLMs already assign non-trivial probability mass to correct multi-step solutions; the bottleneck is locating these…

cs.CL2026

SRA: Span Representation Alignment for Large Language Model Distillation

Quoc Phong Dao, Hoang Son Nguyen, Pham Khanh Chi +4

Cross-Tokenizer Knowledge Distillation (CTKD) enables knowledge transfer between a large language model and a smaller student, even when they employ different tokenizers. While exi…

cs.LG2026

Minimizing Collateral Damage in Activation Steering

Tam Nguyen, Tu Anh Nguyen, Sina Alemohammad +1

Activation steering is a method for controlling Large Language Model (LLM) behavior by intervening in its internal representations to increase the alignment with a specific target…

cs.CL2026

Linguistically Informed Evaluation of Multilingual ASR for African Languages

Fei-Yueh Chen, Lateef Adeleke, C. M. Downey

Word Error Rate (WER) mischaracterizes ASR models' performance for African languages by combining phonological, tone, and other linguistic errors into a single lexical error. By co…

cs.CV2025

Spatiotemporal Tile-based Attention-guided LSTMs for Traffic Video Prediction

Tu Nguyen

This extended abstract describes our solution for the Traffic4Cast Challenge 2019. The task requires modeling both fine-grained (pixel-level) and coarse (region-level) spatial stru…

cs.CL2025

DM-Codec: Distilling Multimodal Representations for Speech Tokenization

Md Mubtasim Ahasan, Md Fahim, Tasnim Mohiuddin +6

Recent advancements in speech-language models have yielded significant improvements in speech tokenization and synthesis. However, effectively mapping the complex, multidimensional…