From the 1 of 12 linked papers with an AI index.
4 papers · 1 filter
Step-Tagging: Toward controlling the generation of Language Reasoning Models through step monitoring
Yannis Belkhiter, Seshu Tirupathi, Giulio Zizzo +1
The paper proposes Step-Tagging, a lightweight classifier that tags reasoning steps generated by language reasoning models in real time, enabling monitoring and early stopping to r…
TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping
Yannis Belkhiter, Seshu Tirupathi, Giulio Zizzo +1
The field of Language Reasoning Models (LRMs) has been very active over the past few years with advances in training and inference techniques enabling LRMs to reason longer, and mo…
Interpreting LLM-as-a-Judge Policies via Verifiable Global Explanations
Jasmina Gajcin, Erik Miehling, Rahul Nair +3
Using LLMs to evaluate text, that is, LLM-as-a-judge, is increasingly being used at scale to augment or even replace human annotations. As such, it is imperative that we understand…
GAF-Guard: An Agentic Framework for Risk Management and Governance in Large Language Models
Seshu Tirupathi, Dhaval Salwala, Elizabeth Daly +1
As Large Language Models (LLMs) continue to be increasingly applied across various domains, their widespread adoption necessitates rigorous monitoring to prevent unintended negativ…