papers

Publications (6)

cs.CL2023

Improving Opinion-based Question Answering Systems Through Label Error Detection and Overwrite

Xiao Yang, Ahmed K. Mohamed, Shashank Jain +6

Label error is a ubiquitous problem in annotated data. Large amounts of label error substantially degrades the quality of deep learning models. Existing methods to tackle the label…

cs.IR2023

A Study on the Efficiency and Generalization of Light Hybrid Retrievers

Man Luo, Shashank Jain, Anchit Gupta +6

Hybrid retrievers can take advantage of both sparse and dense retrievers. Previous hybrid retrievers leverage indexing-heavy dense retrievers. In this work, we study "Is it possibl…

cs.CV2026

GLIMPSE : Real-Time Text Recognition and Contextual Understanding for VQA in Wearables

Akhil Ramachandran, Ankit Arun, Ashish Shenoy +8

Video Large Language Models (Video LLMs) have shown remarkable progress in understanding and reasoning about visual content, particularly in tasks involving text recognition and te…

cs.CV2024

EgoQR: Efficient QR Code Reading in Egocentric Settings

Mohsen Moslehpour, Yichao Lu, Pierce Chuang +7

QR codes have become ubiquitous in daily life, enabling rapid information exchange. With the increasing adoption of smart wearable devices, there is a need for efficient, and frict…

cs.CV2024

Lumos : Empowering Multimodal LLMs with Scene Text Recognition

Ashish Shenoy, Yichao Lu, Srihari Jayakumar +11

We introduce Lumos, the first end-to-end multimodal question-answering system with text understanding capabilities. At the core of Lumos is a Scene Text Recognition (STR) component…

cs.CL2021

Conversational Answer Generation and Factuality for Reading Comprehension Question-Answering

Stan Peshterliev, Barlas Oguz, Debojeet Chatterjee +2

Question answering (QA) is an important use case on voice assistants. A popular approach to QA is extractive reading comprehension (RC) which finds an answer span in a text passage…