Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
Training Noise Token Pruning
Mingxing Rao, Bohan Jiang, Daniel Moyer
In the present work we present Training Noise Token (TNT) Pruning for vision transformers. Our method relaxes the discrete token dropping condition to continuous additive noise, pr…
cs.CV2024
Zero-shot Prompt-based Video Encoder for Surgical Gesture Recognition
Mingxing Rao, Yinhong Qin, Soheil Kolouri +2
Purpose: In order to produce a surgical gesture recognition system that can support a wide variety of procedures, either a very large annotated dataset must be acquired, or fitted…