From the 7 of 188 papers with an AI index.
53 citations
- J. Wang18 profiles100 · h 40
- X. Li11 profiles100 · h 55
- Z. Liu13 profiles89 · h 35
- C. Li15 profiles88 · h 80
- H. Li18 profiles86 · h 91
- Z. Wang17 profiles80 · h 49
- Y. Wang22 profiles79 · h 77
- S. Bhattacharya3 profiles78
- Y. Li14 profiles75 · h 32
- X. Chen8 profiles72 · h 35
- Y. Sun7 profiles72 · h 70
- Z. Li13 profiles70 · h 27
- Tsinghua UniversityCN57 papers
- Institute of Modern PhysicsCN56 papers
- Peking UniversityCN56 papers
- University of Science and Technology of ChinaCN53 papers
- South China Normal UniversityCN48 papers
- Istituto Nazionale di Fisica Nucleare, Laboratori Nazionali di FrascatiIT47 papers
- University of TarapacáCL46 papers
- Zhejiang UniversityCN46 papers
- Beihang UniversityCN44 papers
- National Centre for Nuclear ResearchPL44 papers
- University of TurinIT44 papers
- Carnegie Mellon UniversityUS43 papers
6 papers · 1 filter
Towards Physically Realizable Adversarial Attenuation Patch against SAR Object Detection
Yiming Zhang, Weibo Qin, Feng Wang
Deep neural networks have demonstrated excellent performance in SAR target detection tasks but remain susceptible to adversarial attacks. Existing SAR-specific attack methods can e…
Broken Memories: Detecting and Mitigating Memorization in Diffusion Models with Degraded Generations
Yuanmin Huang, Mi Zhang, Chen Chen +4
While diffusion models excel at generating high-quality images, their tendency to memorize training data poses significant privacy and copyright risks. In this work, we for the fir…
Image-Conditioned Instance Prompt Network for Referring Remote Sensing Image Segmentation
Biaoyu Ren, Qingsheng Wang, Cun Xu +2
Referring Remote Sensing Image Segmentation (RRSIS) is a situated, task-driven cross-modal task related to the embodied perception paradigm, requiring models to align visual-spatia…
When Large Vision-Language Models Meet Person Re-Identification
Qizao Wang, Bin Li, Xiangyang Xue
Large Vision-Language Models (LVLMs) that incorporate visual models and large language models have achieved impressive results across cross-modal understanding and reasoning tasks.…
EchoGen: Cycle-Consistent Learning for Unified Layout-Image Generation and Understanding
Kai Zou, Hongbo Liu, Dian Zheng +3
In this work, we present EchoGen, a unified framework for layout-to-image generation and image grounding, capable of generating images with accurate layouts and high fidelity to te…
Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
Xianqiang Gao, Pingrui Zhang, Delin Qu +4
3D Object Affordance Grounding aims to predict the functional regions on a 3D object and has laid the foundation for a wide range of applications in robotics. Recent advances tackl…