activity
20242026
collaborators

5 papers

cs.LG2026

ASTRA: Communication-Efficient Acceleration for Multi-Device Transformer Inference

Xiao Liu, Lijun Zhang, Deepak Ganesan +1

Multi-device inference can reduce Transformer latency by parallelizing computation. However, existing methods require high inter-device bandwidth, making them impractical for bandw…

cs.CV2026

Aligned Vector Quantization for Edge-Cloud Collabrative Vision-Language Models

Xiao Liu, Lijun Zhang, Deepak Ganesan +1

Vision Language Models (VLMs) are central to Visual Question Answering (VQA) systems and are typically deployed in the cloud due to their high computational demands. However, this…

cs.LG2025

Attacking All Tasks at Once Using Adversarial Examples in Multi-Task Learning

Lijun Zhang, Xiao Liu, Kaleel Mahmood +2

Visual content understanding frequently relies on multi-task models to extract robust representations of a single visual input for multiple downstream tasks. However, in comparison…

cs.LG2025

Reimagining Parameter Space Exploration with Diffusion Models

Lijun Zhang, Xiao Liu, Hui Guan

Adapting neural networks to new tasks typically requires task-specific fine-tuning, which is time-consuming and reliant on labeled data. We explore a generative alternative that pr…

cs.CV2024

Attack-Resilient Image Watermarking Using Stable Diffusion

Lijun Zhang, Xiao Liu, Antoni Viros Martin +3

Watermarking images is critical for tracking image provenance and proving ownership. With the advent of generative models, such as stable diffusion, that can create fake but realis…