3 papers
q-bio.NC2026
BEAST3D: Animal behavioral analysis and neural encoding from multi-view video via Gaussian splatting
Yanchen Wang, Lenny Aharon, Wangshu Zhu +7
Multi-view video recordings are increasingly used to capture the 3D movements of animals in experimental settings, yet extracting rich 3D representations from these recordings rema…
cs.SD2026
Are Audio-Language Models Listening? Audio-Specialist Heads for Adaptive Audio Steering
Neta Glazer, Lenny Aharon, Ethan Fetaya
Multimodal large language models can exhibit text dominance, over-relying on linguistic priors instead of grounding predictions in non-text inputs. One example is large audio-langu…
cs.CV2025
An uncertainty-aware framework for data-efficient multi-view animal pose estimation
Lenny Aharon, Keemin Lee, Karan Sikka +4
Multi-view pose estimation is essential for quantifying animal behavior in scientific research, yet current methods struggle to achieve accurate tracking with limited labeled data…