activity
20242026
most citedCan Large Language Models Infer Causal Relationships from Real-World Text?

2 citations · 5 across the 55 of their papers we have counts for

collaborators
Showing cs.CLShow all

45 papers · 1 filter

cs.CL2026

Khondo: A Multimodal Benchmark for Document Packet Splitting of Bangla Forms

Abu Tyeb Azad, Fahim Ahmed, Ishita Sur Apan +7

Document packets, multiple documents concatenated into a single file, are common in government and administrative workflows, yet splitting them into their constituent documents is…

cs.CL2026

BaFCo: A Document Understanding Benchmark for Complex Bangla Form Comprehension

Abu Tyeb Azad, Ishita Sur Apan, Fahim Ahmed +8

Document comprehension is a challenging yet impactful task for Multimodal Large Language Models, especially as these systems see growing adoption in real-world, human-centric appli…

cs.CL2026

RECOM: A Validity Discrimination Tradeoff in Automatic Metrics for Open Ended Reddit Question Answering

Pushwitha Krishnappa, Amit Das, Vinija Jain +2

Automatic metrics are the default for evaluating LLM-generated text, yet a metric is quietly asked to do two jobs: tell genuine content alignment from surface coincidence (validity…

cs.CL2026

From Concept-Aligned Tokens to Vulnerable Features: Mechanistic Localization of Jailbreaks

Nilanjana Das, Mathew Dawit, Aman Chadha +1

Jailbreak attacks expose a persistent failure mode in safety-aligned LLMs: models can be pushed into harmful behavior, but the internal representations enabling this shift remain p…

cs.CL2026

Neural FOXP2 -- Language Specific Neuron Steering for Targeted Language Improvement in LLMs

Anusa Saha, Tanmay Joshi, Vinija Jain +2

LLMs are multilingual by training, yet their lingua franca is often English, reflecting English language dominance in pretraining. Other languages remain in parametric memory but a…

cs.CL2026

MENTIS: What Belief Changes Under Alignment? Measuring Multi-Scale Latent Torsion in Language Models

Partha Pratim Saha, Samarth Raina, Mayur Parvatikar +4

Preference alignment has substantially improved the observable behavior of large language models, yet it remains unclear what alignment changes internally. Aligned systems still fa…