7 papers
Cognitive Fatigue in Autoregressive Transformers: Formalization and Measurement
Riju Marwah, Ritvik Garimella, Vishal Pallagani +3
Autoregressive language models frequently degrade during long-horizon generation, producing repetitive text, losing instruction adherence, and exhibiting unstable entropy. Despite…
Fast approximate Bayesian multidimensional scaling with consistency guarantees
Ami Sheth, Aaron Smith, Andrew J. Holbrook
Bayesian multidimensional scaling (BMDS) embeds objects in a low-dimensional space to approximately preserve an observed dissimilarity matrix. Compared to classic MDS, BMDS is…
CANDI: Contextual Alignment for Niche Domains Question Answering
Megha Chakraborty, Darssan L. Eswaramoorthi, Het Riteshkumar Shah +5
The deployment of large language models (LLMs) in specialized domains like medical diagnostics and financial advisory necessitates evaluating capabilities beyond general knowledge.…
SAAG: Structured Agent Assessment and Grounding
Ritvik Garimella, Vedant Khandelwal, Anvi Kohli +1
Exact-match evaluation of agent-calling obscures qualitatively different failure modes: a model may select the right function yet hallucinate argument values, or satisfy a schema w…
Chatsparent: An Interactive System for Detecting and Mitigating Cognitive Fatigue in LLMs
Riju Marwah, Vishal Pallagani, Ritvik Garimella +1
LLMs are increasingly being deployed as chatbots, but today's interfaces offer little to no friction: users interact through seamless conversations that conceal when the model is d…
NeuroLit Navigator: A Neurosymbolic Approach to Scholarly Article Searches for Systematic Reviews
Vedant Khandelwal, Kaushik Roy, Valerie Lookingbill +4
The introduction of Large Language Models (LLMs) has significantly impacted various fields, including education, for example, by enabling the creation of personalized learning mate…