Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Hardware-Aware Learned Representation Compression for Distributed In-Sensor Vision
Chengwei Zhou, Abu Masum, Xuming Chen +7
In-sensor computing reduces the cost of transmitting high-resolution image data by performing early-stage processing near the sensor. However, the logic chip integrated with a CMOS…
cs.LG2025
EntroLLM: Entropy Encoded Weight Compression for Efficient Large Language Model Inference on Edge Devices
Arnab Sanyal, Gourav Datta, Prithwish Mukherjee +2
Large Language Models (LLMs) achieve strong performance across tasks, but face storage and compute challenges on edge devices. We propose EntroLLM, a compression framework combinin…