1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CR2026
CanaryBench: Stress Testing Privacy Leakage in Cluster-Level Conversation Summaries
Deep Mehta
Aggregate analytics over conversational data are increasingly used for safety monitoring, governance, and product analysis in large language model systems. A common practice is to…
cs.AI2026★ 1 cited
Does Inference Scaling Improve Reasoning Faithfulness? A Multi-Model Analysis of Self-Consistency Tradeoffs
Deep Mehta
Self-consistency has emerged as a popular technique for improving large language model accuracy on reasoning tasks. The approach is straightforward: generate multiple reasoning pat…