7 citations · 8 across the 2 of their papers we have counts for
3 papers
cs.CV2025★ 1 cited
OmniVinci: Enhancing Architecture and Data for Omni-Modal Understanding LLM
Hanrong Ye, Chao-Han Huck Yang, Arushi Goel +29
Advancing machine intelligence requires developing the ability to perceive across multiple modalities, much as humans sense the world. We introduce OmniVinci, an initiative to buil…
cs.CV2017
Generating Multiple Diverse Hypotheses for Human 3D Pose Consistent with 2D Joint Detections
Ehsan Jahangiri, Alan L. Yuille
We propose a method to generate multiple diverse and valid human pose hypotheses in 3D all consistent with the 2D detection of joints in a monocular RGB image. We use a novel gener…
cs.CV2017★ 7 cited
Information Pursuit: A Bayesian Framework for Sequential Scene Parsing
Ehsan Jahangiri, Erdem Yoruk, Rene Vidal +2
Despite enormous progress in object detection and classification, the problem of incorporating expected contextual relationships among object instances into modern recognition syst…