4 papers
Lie to Me: Knowledge Graphs for Robust Hallucination Self-Detection in LLMs
Sahil Kale, Antonio Luca Alfeo
Hallucinations, the generation of apparently convincing yet false statements, remain a major barrier to the safe deployment of LLMs. Building on the strong performance of self-dete…
TeXpert: A Multi-Level Benchmark for Evaluating LaTeX Code Generation by LLMs
Sahil Kale, Vijaykant Nadadur
LaTeX's precision and flexibility in typesetting have made it the gold standard for the preparation of scientific documentation. Large Language Models (LLMs) present a promising op…
Mirage of Mastery: Memorization Tricks LLMs into Artificially Inflated Self-Knowledge
Sahil Kale
When artificial intelligence mistakes memorization for intelligence, it creates a dangerous mirage of reasoning. Existing studies treat memorization and self-knowledge deficits in…
Line of Duty: Evaluating LLM Self-Knowledge via Consistency in Feasibility Boundaries
Sahil Kale, Vijaykant Nadadur
As LLMs grow more powerful, their most profound achievement may be recognising when to say "I don't know". Existing studies on LLM self-knowledge have been largely constrained by h…