Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
StructFormer: Document Structure-based Masked Attention and its Impact on Language Model Pre-Training
Kaustubh Ponkshe, Venkatapathy Subramanian, Natwar Modani +1
Most state-of-the-art techniques for Language Models (LMs) today rely on transformer-based architectures and their ubiquitous attention mechanism. However, the exponential growth i…
cs.CL2024
GenSco: Can Question Decomposition based Passage Alignment improve Question Answering?
Barah Fazili, Koustava Goswami, Natwar Modani +1
Retrieval augmented generation (RAG) with large language models (LLMs) for Question Answering (QA) entails furnishing relevant context within the prompt to facilitate the LLM in an…