Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Quantifying Ranking Uncertainty in LLM Benchmarks
Bitya Neuhof, Yuval Benjamini
Pretrained models are typically ranked on multi-task leaderboards to assess their effectiveness across diverse tasks. Rank confidence intervals were recently introduced as a method…
cs.LG2025
CardiCat: a Variational Autoencoder for High-Cardinality Tabular Data
Lee Carlin, Yuval Benjamini
High-cardinality categorical features are a common characteristic of mixed-type tabular datasets. Existing generative model architectures struggle to learn the complexities of such…
cs.LG2024
Class Distribution Shifts in Zero-Shot Learning: Learning Robust Representations
Yuli Slavutsky, Yuval Benjamini
Zero-shot learning methods typically assume that the new, unseen classes encountered during deployment come from the same distribution as the the classes in the training set. Howev…