6 papers
From Sparse Features to Trustworthy Proxies: Certifying SAE-Based Interpretability
Dibyanayan Bandyopadhyay, Asif Ekbal
Sparse autoencoders (SAEs) are increasingly used to extract interpretable features from language models (LMs), yet a central question remains: when can an SAE-based explanation be…
Sparse Semantic Dimension as a Generalization Certificate for LLMs
Dibyanayan Bandyopadhyay, Asif Ekbal
Standard statistical learning theory predicts that Large Language Models (LLMs) should overfit because their parameter counts vastly exceed the number of training tokens. Yet, in p…
Bridging the Linguistic Divide: A Survey on Leveraging Large Language Models for Machine Translation
Baban Gain, Dibyanayan Bandyopadhyay, Asif Ekbal +1
Large Language Models (LLMs) are rapidly reshaping machine translation (MT), particularly by introducing instruction-following, in-context learning, and preference-based alignment…
CAuSE: Decoding Multimodal Classifiers using Faithful Natural Language Explanation
Dibyanayan Bandyopadhyay, Soham Bhattacharjee, Mohammed Hasanuzzaman +1
Multimodal classifiers function as opaque black box models. While several techniques exist to interpret their predictions, very few of them are as intuitive and accessible as natur…
Impact of Visual Context on Noisy Multimodal NMT: An Empirical Study for English to Indian Languages
Baban Gain, Dibyanayan Bandyopadhyay, Samrat Mukherjee +2
Neural Machine Translation (NMT) has made remarkable progress using large-scale textual data, but the potential of incorporating multimodal inputs, especially visual information, r…
Thinking Machines: A Survey of LLM based Reasoning Strategies
Dibyanayan Bandyopadhyay, Soham Bhattacharjee, Asif Ekbal
Large Language Models (LLMs) are highly proficient in language-based tasks. Their language capabilities have positioned them at the forefront of the future AGI (Artificial General…