3 papers
cs.CV2026
Propose and Attend: Training-free MLLM Grounding Confidence via Multi-Token Localized Attention
Daniel Shalam, Emanuel Ben Baruch, Avi Ben Cohen +1
Multimodal large language models can emit localized predictions, bounding boxes for objects and temporal windows for video and audio events, but they hallucinate these regions prol…
cs.CL2025
Group-Aware Reinforcement Learning for Output Diversity in Large Language Models
Oron Anschel, Alon Shoshan, Adam Botach +7
Large Language Models (LLMs) often suffer from mode collapse, repeatedly generating the same few completions even when many valid answers exist, limiting their diversity across a w…
cs.CV2023
FPGAN-Control: A Controllable Fingerprint Generator for Training with Synthetic Data
Alon Shoshan, Nadav Bhonker, Emanuel Ben Baruch +5
Training fingerprint recognition models using synthetic data has recently gained increased attention in the biometric community as it alleviates the dependency on sensitive persona…