358 citations · 873 across the 23 of their papers we have counts for
Showing 2025Show all
2 papers · 1 filter
cs.LG2025
Quantization-Free Autoregressive Action Transformer
Ziyad Sheebaelhamd, Michael Tschannen, Michael Muehlebach +1
Current transformer-based imitation learning approaches introduce discrete action representations and train an autoregressive transformer decoder on the resulting latent code. Howe…
cs.CV2025★ 17 cited
SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Michael Tschannen, Alexey Gritsenko, Xiao Wang +11
We introduce SigLIP 2, a family of new multilingual vision-language encoders that build on the success of the original SigLIP. In this second iteration, we extend the original imag…