Publications (18)
Generative Adversarial Networks: A Survey Towards Private and Secure Applications
Zhipeng Cai, Zuobin Xiong, Honghui Xu +3
Generative Adversarial Networks (GAN) have promoted a variety of applications in computer vision, natural language processing, etc. due to its generative model's compelling ability…
AP-GRPO: Anchor-Gated Phonetic Alignment with Policy Optimization for Pathological Speech Reconstruction
Pengfei Zhang, Hoang H Nguyen, Yutong Song +6
Pathological speech from patients with neurodegenerative and neuromotor disorders is often acoustically distorted and linguistically fragmented, making pathological speech reconstr…
Enhancing Object Coherence in Layout-to-Image Synthesis
Yibin Wang, Changhai Zhou, Honghui Xu
Layout-to-image synthesis is an emerging technique in conditional image generation. It aims to generate complex scenes, where users require fine control over the layout of the obje…
MVG-KAN: Multi-View Geo-Wind Guided KAN for PM Forecasting
Cheng Huang, Muyao Guan, Jairus Yougui Railey +6
Accurate short-term PM forecasting is important for public health protection, air-quality early warning, and urban environmental management. However, PM variation i…
DP-FedLoRA: Privacy-Enhanced Federated Fine-Tuning for On-Device Large Language Models
Honghui Xu, Shiva Shrestha, Wei Chen +2
As on-device large language model (LLM) systems become increasingly prevalent, federated fine-tuning enables advanced language understanding and generation directly on edge devices…
A Survey: Towards Privacy and Security in Mobile Large Language Models
Honghui Xu, Kaiyang Li, Wei Chen +3
Mobile Large Language Models (LLMs) are revolutionizing diverse fields such as healthcare, finance, and education with their ability to perform advanced natural language processing…
MedSpeak: A Knowledge Graph-Aided ASR Error Correction Framework for Spoken Medical QA
Yutong Song, Shiva Shrestha, Chenhan Lyu +5
Spoken question-answering (SQA) systems relying on automatic speech recognition (ASR) often struggle with accurately recognizing medical terminology. To this end, we propose MedSpe…
AgriSentinel: Privacy-Enhanced Embedded-LLM Crop Disease Alerting System
Chanti Raju Mylay, Bobin Deng, Zhipeng Cai +1
Crop diseases pose significant threats to global food security, agricultural productivity, and sustainable farming practices, directly affecting farmers' livelihoods and economic s…
DemMA: Dementia Multi-Turn Dialogue Agent with Expert-Guided Reasoning and Action Simulation
Yutong Song, Jiang Wu, Kazi Sharif +3
Simulating dementia patients with large language models (LLMs) is challenging due to the need to jointly model cognitive impairment, emotional dynamics, and nonverbal behaviors ove…
The Robustness of Spiking Neural Networks in Communication and its Application towards Network Efficiency in Federated Learning
Manh V. Nguyen, Liang Zhao, Bobin Deng +3
Spiking Neural Networks (SNNs) have recently gained significant interest in on-chip learning in embedded devices and emerged as an energy-efficient alternative to conventional Arti…
Clustering with Similarity Preserving
Zhao Kang, Honghui Xu, Boyu Wang +2
Graph-based clustering has shown promising performance in many tasks. A key step of graph-based approach is the similarity graph construction. In general, learning graph in kernel…
DreamText: High Fidelity Scene Text Synthesis
Yibin Wang, Weizhong Zhang, Honghui Xu +1
Scene text synthesis involves rendering specified texts onto arbitrary images. Current methods typically formulate this task in an end-to-end manner but lack effective character-le…
GA-HQS: MRI reconstruction via a generically accelerated unfolding approach
Jiawei Jiang, Yuchao Feng, Honghui Xu +2
Deep unfolding networks (DUNs) are the foremost methods in the realm of compressed sensing MRI, as they can employ learnable networks to facilitate interpretable forward-inference…
Security Risks Concerns of Generative AI in the IoT
Honghui Xu, Yingshu Li, Olusesi Balogun +3
In an era where the Internet of Things (IoT) intersects increasingly with generative Artificial Intelligence (AI), this article scrutinizes the emergent security risks inherent in…
From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models
Pengfei Zhang, Hoang H Nguyen, Kazi Shaharair Sharif +6
Recent Large Audio Language Models (LALMs) have achieved remarkable progress in audio perceptual tasks across individual acoustic layers, including speech, sound, and music. Howeve…
When FinTech Meets Privacy: Securing Financial LLMs with Differential Private Fine-Tuning
Sichen Zhu, Hoyeung Leung, Xiaoyi Wang +2
The integration of Large Language Models (LLMs) into financial technology (FinTech) has revolutionized the analysis and processing of complex financial data, driving advancements i…
C3S3: Complementary Competition and Contrastive Selection for Semi-Supervised Medical Image Segmentation
Jiaying He, Yitong Lin, Jiahe Chen +2
For the immanent challenge of insufficiently annotated samples in the medical field, semi-supervised medical image segmentation (SSMIS) offers a promising solution. Despite achievi…
Can the Environment Speak for Itself? -GRPO: A Turn-Trajectory Group Relative Policy Optimization for Caregiver Agents
Yutong Song, Jiang Wu, Pengfei Zhang +4
Optimizing large language models (LLMs) for long-horizon caregiver agents requires balancing delayed task objectives with immediate environment dynamics, such as patient distress a…