works on

From the 1 of 115 linked papers with an AI index.

activity
20242026
most citedDeep Learning-Based Object Pose Estimation: A Comprehensive Survey

21 citations · 23 across the 29 of their papers we have counts for

collaborators
Showing cs.CVShow all

96 papers · 1 filter

cs.CV2026

Evaluating Newtonian Mechanics in Video Generative Models with Real Physical Systems

Antonios Tragoudaras, Chenyu Zhang, Daniil Cherniavskii +7

Recent advances in image and video generation raise hopes that these models possess world modeling capabilities-the ability to generate realistic, physically plausible videos. This…

cs.CV2026

Chorus: Multi-Teacher Pretraining for Holistic 3D Gaussian Scene Encoding

Yue Li, Qi Ma, Runyi Yang +8

While 3DGS has emerged as a high-fidelity scene representation, encoding rich, general-purpose features directly from its primitives remains under-explored. We address this gap by…

cs.CV2026

Hallucination Early Detection in Diffusion Models

Federico Betti, Lorenzo Baraldi, Rita Cucchiara +1

Text-to-Image generation has seen significant advancements in output realism with the advent of diffusion models. However, diffusion models encounter difficulties when tasked with…

cs.CV2026

NullFace: Training-Free Localized Face Anonymization

Han-Wei Kung, Tuomas Varanka, Terence Sim +1

Privacy concerns around ever increasing number of cameras are increasing in today's digital age. Although existing anonymization methods are able to obscure identity information, t…

cs.CV2026

The Second Challenge on Cross-Domain Few-Shot Object Detection at NTIRE 2026: Methods and Results

Xingyu Qiu, Yuqian Fu, Jiawei Geng +70

Cross-domain few-shot object detection (CD-FSOD) remains a challenging problem for existing object detectors and few-shot learning approaches, particularly when generalizing across…

cs.CV2026

Finetune Like You Pretrain: Boosting Zero-shot Adversarial Robustness in Vision-language Models

Songlong Xing, Weijie Wang, Zhengyu Zhao +3

Despite their impressive zero-shot abilities, vision-language models such as CLIP have been shown to be susceptible to adversarial attacks. To enhance its adversarial robustness, r…