activity
20222026
most citedYETI (YET to Intervene) Proactive Interventions by Multimodal AI Agents in Augmented Reality Tasks

1 citations · 1 across the 4 of their papers we have counts for

collaborators

5 papers

cs.CV2026

From Videos to Conversations: Egocentric Instructions for Task Assistance

Lavisha Aggarwal, Vikas Bahirwani, Andrea Colaco

Many everyday tasks, ranging from appliance repair and cooking to car maintenance, require expert knowledge, particularly for complex, multi-step procedures. Despite growing intere…

cs.CV2025

Generating Dialogues from Egocentric Instructional Videos for Task Assistance: Dataset, Method and Benchmark

Lavisha Aggarwal, Vikas Bahirwani, Lin Li +1

Many everyday tasks ranging from fixing appliances, cooking recipes to car maintenance require expert knowledge, especially when tasks are complex and multi-step. Despite growing i…

cs.CV2025

EgoTrigger: Toward Audio-Driven Image Capture for Human Memory Enhancement in All-Day Energy-Efficient Smart Glasses

Akshay Paruchuri, Sinan Hersek, Lavisha Aggarwal +6

All-day smart glasses are likely to emerge as platforms capable of continuous contextual sensing, uniquely positioning them for unprecedented assistance in our daily lives. Integra…

cs.AI20251 cited

YETI (YET to Intervene) Proactive Interventions by Multimodal AI Agents in Augmented Reality Tasks

Saptarashmi Bandyopadhyay, Vikas Bahirwani, Lavisha Aggarwal +3

Multimodal AI Agents are AI models that have the capability of interactively and cooperatively assisting human users to solve day-to-day tasks. Augmented Reality (AR) head worn dev…

cs.CV2022

Identity Preserving Loss for Learned Image Compression

Jiuhong Xiao, Lavisha Aggarwal, Prithviraj Banerjee +2

Deep learning model inference on embedded devices is challenging due to the limited availability of computation resources. A popular alternative is to perform model inference on th…