papers

Publications (18)

cs.LG2021

Generative Adversarial Networks: A Survey Towards Private and Secure Applications

Zhipeng Cai, Zuobin Xiong, Honghui Xu +3

Generative Adversarial Networks (GAN) have promoted a variety of applications in computer vision, natural language processing, etc. due to its generative model's compelling ability…

cs.SD2026

AP-GRPO: Anchor-Gated Phonetic Alignment with Policy Optimization for Pathological Speech Reconstruction

Pengfei Zhang, Hoang H Nguyen, Yutong Song +6

Pathological speech from patients with neurodegenerative and neuromotor disorders is often acoustically distorted and linguistically fragmented, making pathological speech reconstr…

cs.CV2025

Enhancing Object Coherence in Layout-to-Image Synthesis

Yibin Wang, Changhai Zhou, Honghui Xu

Layout-to-image synthesis is an emerging technique in conditional image generation. It aims to generate complex scenes, where users require fine control over the layout of the obje…

cs.AI2026

MVG-KAN: Multi-View Geo-Wind Guided KAN for PM Forecasting

Cheng Huang, Muyao Guan, Jairus Yougui Railey +6

Accurate short-term PM forecasting is important for public health protection, air-quality early warning, and urban environmental management. However, PM variation i…

cs.CR2025

DP-FedLoRA: Privacy-Enhanced Federated Fine-Tuning for On-Device Large Language Models

Honghui Xu, Shiva Shrestha, Wei Chen +2

As on-device large language model (LLM) systems become increasingly prevalent, federated fine-tuning enables advanced language understanding and generation directly on edge devices…

cs.CR2025

A Survey: Towards Privacy and Security in Mobile Large Language Models

Honghui Xu, Kaiyang Li, Wei Chen +3

Mobile Large Language Models (LLMs) are revolutionizing diverse fields such as healthcare, finance, and education with their ability to perform advanced natural language processing…

cs.CL2026

MedSpeak: A Knowledge Graph-Aided ASR Error Correction Framework for Spoken Medical QA

Yutong Song, Shiva Shrestha, Chenhan Lyu +5

Spoken question-answering (SQA) systems relying on automatic speech recognition (ASR) often struggle with accurately recognizing medical terminology. To this end, we propose MedSpe…

cs.CR2025

AgriSentinel: Privacy-Enhanced Embedded-LLM Crop Disease Alerting System

Chanti Raju Mylay, Bobin Deng, Zhipeng Cai +1

Crop diseases pose significant threats to global food security, agricultural productivity, and sustainable farming practices, directly affecting farmers' livelihoods and economic s…

cs.MA2026

DemMA: Dementia Multi-Turn Dialogue Agent with Expert-Guided Reasoning and Action Simulation

Yutong Song, Jiang Wu, Kazi Sharif +3

Simulating dementia patients with large language models (LLMs) is challenging due to the need to jointly model cognitive impairment, emotional dynamics, and nonverbal behaviors ove…

cs.LG2024

The Robustness of Spiking Neural Networks in Communication and its Application towards Network Efficiency in Federated Learning

Manh V. Nguyen, Liang Zhao, Bobin Deng +3

Spiking Neural Networks (SNNs) have recently gained significant interest in on-chip learning in embedded devices and emerged as an energy-efficient alternative to conventional Arti…

cs.LG2019

Clustering with Similarity Preserving

Zhao Kang, Honghui Xu, Boyu Wang +2

Graph-based clustering has shown promising performance in many tasks. A key step of graph-based approach is the similarity graph construction. In general, learning graph in kernel…

cs.CV2025

DreamText: High Fidelity Scene Text Synthesis

Yibin Wang, Weizhong Zhang, Honghui Xu +1

Scene text synthesis involves rendering specified texts onto arbitrary images. Current methods typically formulate this task in an end-to-end manner but lack effective character-le…

eess.IV2023

GA-HQS: MRI reconstruction via a generically accelerated unfolding approach

Jiawei Jiang, Yuchao Feng, Honghui Xu +2

Deep unfolding networks (DUNs) are the foremost methods in the realm of compressed sensing MRI, as they can employ learnable networks to facilitate interpretable forward-inference…

cs.CR2024

Security Risks Concerns of Generative AI in the IoT

Honghui Xu, Yingshu Li, Olusesi Balogun +3

In an era where the Internet of Things (IoT) intersects increasingly with generative Artificial Intelligence (AI), this article scrutinizes the emergent security risks inherent in…

cs.SD2026

From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models

Pengfei Zhang, Hoang H Nguyen, Kazi Shaharair Sharif +6

Recent Large Audio Language Models (LALMs) have achieved remarkable progress in audio perceptual tasks across individual acoustic layers, including speech, sound, and music. Howeve…

cs.CR2025

When FinTech Meets Privacy: Securing Financial LLMs with Differential Private Fine-Tuning

Sichen Zhu, Hoyeung Leung, Xiaoyi Wang +2

The integration of Large Language Models (LLMs) into financial technology (FinTech) has revolutionized the analysis and processing of complex financial data, driving advancements i…

cs.CV2025

C3S3: Complementary Competition and Contrastive Selection for Semi-Supervised Medical Image Segmentation

Jiaying He, Yitong Lin, Jiahe Chen +2

For the immanent challenge of insufficiently annotated samples in the medical field, semi-supervised medical image segmentation (SSMIS) offers a promising solution. Despite achievi…

cs.AI2026

Can the Environment Speak for Itself? -GRPO: A Turn-Trajectory Group Relative Policy Optimization for Caregiver Agents

Yutong Song, Jiang Wu, Pengfei Zhang +4

Optimizing large language models (LLMs) for long-horizon caregiver agents requires balancing delayed task objectives with immediate environment dynamics, such as patient distress a…