2 papers
cs.CL2024
StructFormer: Document Structure-based Masked Attention and its Impact on Language Model Pre-Training
Kaustubh Ponkshe, Venkatapathy Subramanian, Natwar Modani +1
Most state-of-the-art techniques for Language Models (LMs) today rely on transformer-based architectures and their ubiquitous attention mechanism. However, the exponential growth i…
cs.CV2024
TEXTRON: Weakly Supervised Multilingual Text Detection through Data Programming
Dhruv Kudale, Badri Vishal Kasuba, Venkatapathy Subramanian +2
Several recent deep learning (DL) based techniques perform considerably well on image-based multilingual text detection. However, their performance relies heavily on the availabili…