3 papers
cs.CL2025
Assortment of Attention Heads: Accelerating Federated PEFT with Head Pruning and Strategic Client Selection
Yeshwanth Venkatesha, Souvik Kundu, Priyadarshini Panda
Parameter Efficient Fine-Tuning (PEFT) has become the de-facto approach in adapting Large Language Models (LLMs) for downstream tasks in Natural Language Processing. However, its a…
cs.RO2025
Fast and Cost-effective Speculative Edge-Cloud Decoding with Early Exits
Yeshwanth Venkatesha, Souvik Kundu, Priyadarshini Panda
Large Language Models (LLMs) enable various applications on edge devices such as smartphones, wearables, and embodied robots. However, their deployment often depends on expensive c…
cs.RO2025
Intelligent Sensing-to-Action for Robust Autonomy at the Edge: Opportunities and Challenges
Amit Ranjan Trivedi, Sina Tayebati, Hemant Kumawat +9
Autonomous edge computing in robotics, smart cities, and autonomous vehicles relies on the seamless integration of sensing, processing, and actuation for real-time decision-making…