4 papers
Source-free Video Domain Adaptation by Learning from Noisy Labels
Avijit Dasgupta, C. V. Jawahar, Karteek Alahari
Despite the progress seen in classification methods, current approaches for handling videos with distribution shifts in source and target domains remain source-dependent as they re…
Reading Between the Lanes: Text VideoQA on the Road
George Tom, Minesh Mathew, Sergi Garcia +2
Text and signs around roads provide crucial information for drivers, vital for safe navigation and situational awareness. Scene text recognition in motion is a challenging problem,…
NoTeS-Bank: Benchmarking Neural Transcription and Search for Scientific Notes Understanding
Aniket Pal, Sanket Biswas, Alloy Das +6
Understanding and reasoning over academic handwritten notes remains a challenge in document AI, particularly for mathematical equations, diagrams, and scientific notations. Existin…
Towards Deployable OCR models for Indic languages
Minesh Mathew, Ajoy Mondal, CV Jawahar
Recognition of text on word or line images, without the need for sub-word segmentation has become the mainstream of research and development of text recognition for Indian language…