8 papers
How to Verify Consistency of Probabilistic Claims
Orr Paradise, Oliver Richardson, Yoshua Bengio +1
When a probabilistic predictor answers many conditional-probability queries, are its answers self-consistent, and can this be verified in polynomial time? This problem is of intere…
Learning Randomized Reductions
Ferhat Erata, Orr Paradise, Thanos Typaldos +4
Randomized self-reductions (RSRs) express using evaluated at random correlated points, enabling self-correcting programs, instance-hiding protocols, and applications in…
Investigating the Development of Task-Oriented Communication in Vision-Language Models
Boaz Carmeli, Orr Paradise, Shafi Goldwasser +2
We investigate whether \emph{LLM-based agents} can develop task-oriented communication protocols that differ from standard natural language in collaborative reasoning tasks. Our fo…
Models That Prove Their Own Correctness
Noga Amit, Shafi Goldwasser, Orr Paradise +1
How can we trust the correctness of a learned model on a particular input of interest? Model accuracy is typically measured on average over a distribution of inputs, giving no guar…
WhAM: Towards A Translative Model of Sperm Whale Vocalization
Orr Paradise, Pranav Muralikrishnan, Liangyuan Chen +6
Sperm whales communicate in short sequences of clicks known as codas. We present WhAM (Whale Acoustics Model), the first transformer-based model capable of generating synthetic spe…
On Non-interactive Evaluation of Animal Communication Translators
Orr Paradise, David F. Gruber, Adam Tauman Kalai
If you had an AI Whale-to-English translator, how could you validate whether or not it is working? Does one need to interact with the animals or rely on grounded observations such…