4 papers
SIRIN: A Unified Toolkit for Detecting Contextual Hallucinations in Retrieval-Augmented and Memory-Grounded LLM Systems
Julia Belikova, Rauf Parchiev, Mikhail Filimonov +3
SIRIN (Semantic Inconsistency Recognition and Inspection Nexus) is a unified toolkit and interactive web UI for detecting contextual hallucinations (fluent, plausible responses uns…
Hallucination Detection in LLMs with Topological Divergence on Attention Graphs
Alexandra Bazarova, Andrei Volodichev, Aleksandr Yugay +10
Hallucination, i.e., generating factually incorrect content, remains a critical challenge for large language models (LLMs). We introduce TOHA, a TOpology-based HAllucination detect…
Learning Transactions Representations for Information Management in Banks: Mastering Local, Global, and External Knowledge
Alexandra Bazarova, Maria Kovaleva, Ilya Kuleshov +7
In today's world, banks use artificial intelligence to optimize diverse business processes, aiming to improve customer experience. Most of the customer-related tasks can be categor…
From Variability to Stability: Advancing RecSys Benchmarking Practices
Valeriy Shevchenko, Nikita Belousov, Alexey Vasilev +6
In the rapidly evolving domain of Recommender Systems (RecSys), new algorithms frequently claim state-of-the-art performance based on evaluations over a limited set of arbitrarily…