Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
The MASK Benchmark: Disentangling Honesty From Accuracy in AI Systems
Richard Ren, Arunim Agarwal, Mantas Mazeika +13
As large language models (LLMs) become more capable and agentic, the requirement for trust in their outputs grows significantly, yet at the same time concerns have been mounting th…
cs.LG2025
ICEPool: Enhancing Graph Pooling Networks with Inter-cluster Connectivity
Michael Yang
Hierarchical Pooling Models have demonstrated strong performance in classifying graph-structured data. While numerous innovative methods have been proposed to design cluster assign…