activity
20142024
most citedVideoSET: Video Summary Evaluation through Text

42 citations · 72 across the 13 of their papers we have counts for

collaborators

13 papers

cs.CV2024

Multi-Human Mesh Recovery with Transformers

Zeyu Wang, Zhenzhen Weng, Serena Yeung-Levy

Conventional approaches to human mesh recovery predominantly employ a region-based strategy. This involves initially cropping out a human-centered region as a preprocessing step, w…

cs.CV2024

AdaEmbed: Semi-supervised Domain Adaptation in the Embedding Space

Ali Mottaghi, Mohammad Abdullah Jamal, Serena Yeung +1

Semi-supervised domain adaptation (SSDA) presents a critical hurdle in computer vision, especially given the frequent scarcity of labeled data in real-world settings. This scarcity…

cs.LG20239 cited

INSPECT: A Multimodal Dataset for Pulmonary Embolism Diagnosis and Prognosis

Shih-Cheng Huang, Zepeng Huo, Ethan Steinberg +6

Synthesizing information from multiple data sources plays a crucial role in the practice of modern medicine. Current applications of artificial intelligence in medicine often focus…

cs.LG2023

Generalizable Neural Fields as Partially Observed Neural Processes

Jeffrey Gu, Kuan-Chieh Wang, Serena Yeung

Neural fields, which represent signals as a function parameterized by a neural network, are a promising alternative to traditional discrete vector or grid-based representations. Co…

cs.CL2023

Beyond Positive Scaling: How Negation Impacts Scaling Trends of Language Models

Yuhui Zhang, Michihiro Yasunaga, Zhengping Zhou +4

Language models have been shown to exhibit positive scaling, where performance improves as models are scaled up in terms of size, compute, or data. In this work, we introduce NeQA,…

cs.CV20234 cited

ZeroAvatar: Zero-shot 3D Avatar Generation from a Single Image

Zhenzhen Weng, Zeyu Wang, Serena Yeung

Recent advancements in text-to-image generation have enabled significant progress in zero-shot 3D shape generation. This is achieved by score distillation, a methodology that uses…