6 papers
Wiring the 'Why': A Unified Taxonomy and Survey of Abductive Reasoning in LLMs
Moein Salimi, Shaygan Adim, Danial Parnian +3
Regardless of its foundational role in human discovery and sense-making, abductive reasoning--the inference of the most plausible explanation for an observation--has been relativel…
Generalization and Membership Inference Attack a Practical Perspective
Fateme Rahmani, Mahdi Jafari Siavoshani, Mohammad Hossein Rohban
With the emergence of new evaluation metrics and attack methodologies for Membership Inference Attacks (MIA), it becomes essential to reevaluate previously accepted assumptions. In…
Debate as Reward: A Multi-Agent Reward System for Scientific Ideation via RL Post-Training
Moein Salimi, Babak Hosseini Mohtasham, Amin Aghakasiri +6
Large Language Models (LLMs) have demonstrated potential in automating scientific ideation, yet current approaches relying on iterative prompting or complex multi-agent architectur…
Erasure or Erosion? Evaluating Compositional Degradation in Unlearned Text-To-Image Diffusion Models
Arian Komaei Koma, Seyed Amir Kasaei, Ali Aghayari +2
Post-hoc unlearning has emerged as a practical mechanism for removing undesirable concepts from large text-to-image diffusion models. However, prior work primarily evaluates unlear…
Hidden Meanings in Plain Sight: RebusBench for Evaluating Cognitive Visual Reasoning
Seyed Amir Kasaei, Arash Marioriyad, Mahbod Khaleti +3
Large Vision-Language Models (LVLMs) have achieved remarkable proficiency in explicit visual recognition, effectively describing what is directly visible in an image. However, a cr…
Lying to Win: Assessing LLM Deception through Human-AI Games and Parallel-World Probing
Arash Marioriyad, Ali Nouri, Mohammad Hossein Rohban +1
As Large Language Models (LLMs) transition into autonomous agentic roles, the risk of deception-defined behaviorally as the systematic provision of false information to satisfy ext…