1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.CL2025
ROVER: Recursive Reasoning Over Videos with Vision-Language Models for Embodied Tasks
Philip Schroeder, Ondrej Biza, Thomas Weng +2
Vision-language models (VLMs) have exhibited impressive capabilities across diverse image understanding tasks, but still struggle in settings that require reasoning over extended s…
cs.CL2024★ 1 cited
Addition is All You Need for Energy-efficient Language Models
Hongyin Luo, Wei Sun
Large neural networks spend most computation on floating point tensor multiplications. In this work, we find that a floating point multiplier can be approximated by one integer add…