5 citations · 5 across the 3 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Fast SceneScript: Fast and Accurate Language-Based 3D Scene Understanding via Multi-Token Prediction
Ruihong Yin, Xuepeng Shi, Oleksandr Bailo +2
Recent perception-generalist approaches based on language models have achieved state-of-the-art results across diverse tasks, including 3D scene layout estimation and 3D object det…
cs.CV2021
United We Learn Better: Harvesting Learning Improvements From Class Hierarchies Across Tasks
Sindi Shkodrani, Yu Wang, Marco Manfredi +1
Attempts of learning from hierarchical taxonomies in computer vision have been mostly focusing on image classification. Though ways of best harvesting learning improvements from hi…
cs.CV2020★ 5 cited
Shift Equivariance in Object Detection
Marco Manfredi, Yu Wang
Robustness to small image translations is a highly desirable property for object detectors. However, recent works have shown that CNN-based classifiers are not shift invariant. It…