3 papers
cs.CL2025
Persistent Instability in LLM's Personality Measurements: Effects of Scale, Reasoning, and Conversation History
Tommaso Tosato, Saskia Helbling, Yorguin-Jose Mantilla-Ramos +5
Large language models require consistent behavioral patterns for safe deployment, yet there are indications of large variability that may lead to an instable expression of personal…
cs.AI2025
MAFA: A multi-agent framework for annotation
Mahmood Hegazy, Aaron Rodrigues, Azzam Naeem
Modern consumer banking applications require accurate and efficient retrieval of information in response to user queries. Mapping user utterances to the most relevant Frequently As…
cs.IR2025
Enhancing and Scaling Search Query Datasets for Recommendation Systems
Aaron Rodrigues, Mahmood Hegazy, Azzam Naeem
This paper presents a deployed, production-grade system designed to enhance and scale search query datasets for intent-based recommendation systems in digital banking. In real-worl…