4 papers
ByteRover: Agent-Native Memory Through LLM-Curated Hierarchical Context
Andy Nguyen, Danh Doan, Hoang Pham +8
Memory-Augmented Generation (MAG) extends large language models with external memory to support long-context reasoning, but existing approaches universally treat memory as an exter…
Toward Reliable Evaluation of LLM-Based Financial Multi-Agent Systems: Taxonomy, Coordination Primacy, and Cost Awareness
Phat Nguyen, Thang Pham
Multi-agent systems based on large language models (LLMs) for financial trading have grown rapidly since 2023, yet the field lacks a shared framework for understanding what drives…
Token Compression Meets Compact Vision Transformers: A Survey and Comparative Evaluation for Edge AI
Phat Nguyen, Ngai-Man Cheung
Token compression techniques have recently emerged as powerful tools for accelerating Vision Transformer (ViT) inference in computer vision. Due to the quadratic computational comp…
SlimLM: An Efficient Small Language Model for On-Device Document Assistance
Thang M. Pham, Phat T. Nguyen, Seunghyun Yoon +3
While small language models (SLMs) show promises for mobile deployment, their real-world performance and applications on smartphones remains underexplored. We present SlimLM, a ser…