activity
20182026
most citedSuperposition of many models into one

46 citations · 80 across the 23 of their papers we have counts for

collaborators
Showing cs.CVShow all

11 papers · 1 filter

cs.CV2026

MotionPyramid: Hierarchical Motion Representation and Residual Interfaces

Gao Zhu, Zaishuo Xia, Yubei Chen

We ask whether the representational hierarchy seen in perception, from local primitives such as edges to higher level structures such as parts and objects, can be established for m…

cs.CV2025

Scaling Non-Parametric Sampling with Representation

Vincent Lu, Aaron Truong, Zeyu Yun +1

Scaling and architectural advances have produced strikingly photorealistic image generative models, yet their mechanisms still remain opaque. Rather than advancing scaling, our goa…

cs.CV20251 cited

Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMs

Chun-Hsiao Yeh, Chenyu Wang, Shengbang Tong +7

Multi-view understanding, the ability to reconcile visual information across diverse viewpoints for effective navigation, manipulation, and 3D scene comprehension, is a fundamental…

cs.CV2024

Pose-Aware Self-Supervised Learning with Viewpoint Trajectory Regularization

Jiayun Wang, Yubei Chen, Stella X. Yu

Learning visual features from unlabeled images has proven successful for semantic categorization, often by mapping different of the same object to the same feature to achie…

cs.CV2024

Gen4Gen: Generative Data Pipeline for Generative Multi-Concept Composition

Chun-Hsiao Yeh, Ta-Ying Cheng, He-Yen Hsieh +6

Recent text-to-image diffusion models are able to learn and synthesize images containing novel, personalized concepts (e.g., their own pets or specific items) with just a few examp…

cs.CV2023

URLOST: Unsupervised Representation Learning without Stationarity or Topology

Zeyu Yun, Juexiao Zhang, Yann LeCun +1

Unsupervised representation learning has seen tremendous progress. However, it is constrained by its reliance on domain specific stationarity and topology, a limitation not found i…