1 citations · 2 across the 6 of their papers we have counts for
Showing 2025Show all
2 papers · 1 filter
cs.LG2025
TELL-TALE: Task Efficient LLMs with Task Aware Layer Elimination
Omar Naim, Krish Sharma, Niyar R Barman +1
Large Language Models (LLMs) typically come with a fixed architecture, despite growing evidence that not all layers contribute equally to every downstream task. We introduce TALE (…
cs.CL2025
DIMSUM: Discourse in Mathematical Reasoning as a Supervision Module
Krish Sharma, Niyar R Barman, Akshay Chaturvedi +1
We look at reasoning on GSM8k, a dataset of short texts presenting primary school, math problems. We find, with Mirzadeh et al. (2024), that current LLM progress on the data set ma…