5 papers
ASTRA: Communication-Efficient Acceleration for Multi-Device Transformer Inference
Xiao Liu, Lijun Zhang, Deepak Ganesan +1
Multi-device inference can reduce Transformer latency by parallelizing computation. However, existing methods require high inter-device bandwidth, making them impractical for bandw…
Aligned Vector Quantization for Edge-Cloud Collabrative Vision-Language Models
Xiao Liu, Lijun Zhang, Deepak Ganesan +1
Vision Language Models (VLMs) are central to Visual Question Answering (VQA) systems and are typically deployed in the cloud due to their high computational demands. However, this…
Attacking All Tasks at Once Using Adversarial Examples in Multi-Task Learning
Lijun Zhang, Xiao Liu, Kaleel Mahmood +2
Visual content understanding frequently relies on multi-task models to extract robust representations of a single visual input for multiple downstream tasks. However, in comparison…
Reimagining Parameter Space Exploration with Diffusion Models
Lijun Zhang, Xiao Liu, Hui Guan
Adapting neural networks to new tasks typically requires task-specific fine-tuning, which is time-consuming and reliant on labeled data. We explore a generative alternative that pr…
Attack-Resilient Image Watermarking Using Stable Diffusion
Lijun Zhang, Xiao Liu, Antoni Viros Martin +3
Watermarking images is critical for tracking image provenance and proving ownership. With the advent of generative models, such as stable diffusion, that can create fake but realis…