42 citations · 134 across the 11 of their papers we have counts for
15 papers
Where a Strong Backbone Meets Strong Features -- ActionFormer for Ego4D Moment Queries Challenge
Fangzhou Mu, Sicheng Mo, Gillian Wang +1
This report describes our submission to the Ego4D Moment Queries Challenge 2022. Our submission builds on ActionFormer, the state-of-the-art backbone for temporal action localizati…
A Simple Transformer-Based Model for Ego4D Natural Language Queries Challenge
Sicheng Mo, Fangzhou Mu, Yin Li
This report describes Badgers@UW-Madison, our submission to the Ego4D Natural Language Queries (NLQ) Challenge. Our solution inherits the point-based event representation from our…
3D Scene Inference from Transient Histograms
Sacha Jungerman, Atul Ingle, Yin Li +1
Time-resolved image sensors that capture light at pico-to-nanosecond timescales were once limited to niche applications but are now rapidly becoming mainstream in consumer devices.…
mRI: Multi-modal 3D Human Pose Estimation Dataset using mmWave, RGB-D, and Inertial Sensors
Sizhe An, Yin Li, Umit Ogras
The ability to estimate 3D human body pose and movement, also known as human pose estimation (HPE), enables many applications for home-based health monitoring, such as remote rehab…
Learning to Generate Scene Graph from Natural Language Supervision
Yiwu Zhong, Jing Shi, Jianwei Yang +2
Learning from image-text data has demonstrated recent success for many recognition tasks, yet is currently limited to visual features or individual visual concepts such as objects.…
Weakly Supervised Foreground Learning for Weakly Supervised Localization and Detection
Chen-Lin Zhang, Yin Li, Jianxin Wu
Modern deep learning models require large amounts of accurately annotated data, which is often difficult to satisfy. Hence, weakly supervised tasks, including weakly supervised obj…