3 papers
cs.CV2026
Preserve, Then Resolve: Many-to-Many Association and Robust Estimation with General-Purpose Visual Features
Haodong Jiang, Mingzhe Li, Junfeng Wu
The semantic transferability of general-purpose visual features does not guarantee geometric consistency across images. Using frozen DINOv3 features, we show that geometrically cor…
cs.CV2025
EDITOR: Effective and Interpretable Prompt Inversion for Text-to-Image Diffusion Models
Mingzhe Li, Kejing Xia, Gehao Zhang +5
Text-to-image generation models~(e.g., Stable Diffusion) have achieved significant advancements, enabling the creation of high-quality and realistic images based on textual descrip…
cs.CV2024
Real-Time Human Action Recognition on Embedded Platforms
Ruiqi Wang, Zichen Wang, Peiqi Gao +7
With advancements in computer vision and deep learning, video-based human action recognition (HAR) has become practical. However, due to the complexity of the computation pipeline,…