From the 1 of 4 linked papers with an AI index.
4 papers
Not Worth Another Token: Marginal Value Estimation for Efficient Deep Research Agents
Harshitha Kolukuluru, Reshma Ashok, Kirat Arora +7
Long-horizon research agents solve open-ended tasks through iterative retrieval, aggregation, and synthesis, but context grows rapidly while the marginal value of additional eviden…
When Rubrics Change: Cross-Rubric Generalization for Critical Thinking Essay Scoring
Nischal Ashok Kumar, Payu Wittawatolarn, Sana Kang +7
The paper investigates how to train automated essay scoring models on one set of rubrics and apply them to previously unseen rubrics, using fine‑tuned large language models with in…
A Multi-Agent Approach to Validate and Refine LLM-Generated Personalized Math Problems
Fareya Ikram, Nischal Ashok Kumar, Junyang Lu +4
Students benefit from math problems contextualized to their interests. Large language models (LLMs) offer promise for efficient personalization at scale. However, LLM-generated per…
Toward LLM-Supported Automated Assessment of Critical Thinking Subskills
Marisa C. Peczuh, Nischal Ashok Kumar, Ryan Baker +11
As the world becomes increasingly saturated with AI-generated content, disinformation, and algorithmic persuasion, critical thinking - the capacity to evaluate evidence, detect unr…