Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
ARDuP: Active Region Video Diffusion for Universal Policies
Shuaiyi Huang, Mara Levy, Zhenyu Jiang +5
Sequential decision-making can be formulated as a text-conditioned video generation problem, where a video planner, guided by a text-defined goal, generates future frames visualizi…
cs.CV2024
V-VIPE: Variational View Invariant Pose Embedding
Mara Levy, Abhinav Shrivastava
Learning to represent three dimensional (3D) human pose given a two dimensional (2D) image of a person, is a challenging problem. In order to make the problem less ambiguous it has…