6 papers · 1 filter
TMUAD: Enhancing Logical Capabilities in Unified Anomaly Detection Models with a Text Memory Bank
Jiawei Liu, Jiahe Hou, Wei Wang +3
Anomaly detection, which aims to identify anomalies deviating from normal patterns, is challenging due to the limited amount of normal data available. Unlike most existing unified…
SAGA: Surface-Aligned Gaussian Avatar
Ronghan Chen, Yang Cong, Jiayue Liu
This paper presents a Surface-Aligned Gaussian representation for creating animatable human avatars from monocular videos,aiming at improving the novel view and pose synthesis perf…
Learning Generalizable 3D Manipulation With 10 Demonstrations
Yu Ren, Yang Cong, Ronghan Chen +1
Learning robust and generalizable manipulation skills from demonstrations remains a key challenge in robotics, with broad applications in industrial automation and service robotics…
MuseumMaker: Continual Style Customization without Catastrophic Forgetting
Chenxi Liu, Gan Sun, Wenqi Liang +3
Pre-trained large text-to-image (T2I) models with an appropriate text prompt has attracted growing interests in customized images generation field. However, catastrophic forgetting…
Marrying NeRF with Feature Matching for One-step Pose Estimation
Ronghan Chen, Yang Cong, Yu Ren
Given the image collection of an object, we aim at building a real-time image-based pose estimation method, which requires neither its CAD model nor hours of object-specific traini…
Create Your World: Lifelong Text-to-Image Diffusion
Gan Sun, Wenqi Liang, Jiahua Dong +3
Text-to-image generative models can produce diverse high-quality images of concepts with a text prompt, which have demonstrated excellent ability in image generation, image transla…