activity
20242026
collaborators

5 papers

cs.RO2026

STEP: Warm-Started Visuomotor Policies with Spatiotemporal Consistency Prediction

Jinhao Li, Yuxuan Cong, Yingqiao Wang +5

Diffusion policies have recently emerged as a powerful paradigm for visuomotor control in robotic manipulation due to their ability to model the distribution of action sequences an…

cs.AR2026

ROMA: a Read-Only-Memory-based Accelerator for QLoRA-based On-Device LLM

Wenqiang Wang, Yijia Zhang, Zikai Zhang +4

As large language models (LLMs) demonstrate powerful capabilities, deploying them on edge devices has become increasingly crucial, offering advantages in privacy and real-time inte…

cs.AR2026

TOM: A Ternary Read-only Memory Accelerator for LLM-powered Edge Intelligence

Hongyi Guan, Yijia Zhang, Wenqiang Wang +4

The deployment of Large Language Models (LLMs) for real-time intelligence on edge devices is rapidly growing. However, conventional hardware architectures face a fundamental memory…

cs.AR2025

CCSS: Hardware-Accelerated RTL Simulation with Fast Combinational Logic Computing and Sequential Logic Synchronization

Weigang Feng, Yijia Zhang, Zekun Wang +4

As transistor counts in a single chip exceed tens of billions, the complexity of RTL-level simulation and verification has grown exponentially, often extending simulation campaigns…

cs.PF2024

Automating Energy-Efficient GPU Kernel Generation: A Fast Search-Based Compilation Approach

Yijia Zhang, Zhihong Gou, Shijie Cao +4

Deep Neural Networks (DNNs) have revolutionized various fields, but their deployment on GPUs often leads to significant energy consumption. Unlike existing methods for reducing GPU…