From the 1 of 7.9k papers with an AI index.
7.9k citations
- Chinese Academy of SciencesCN1.3k papers
- University of Chinese Academy of SciencesCN1.3k papers
- Kavli Institute for Theoretical SciencesCN1.3k papers
- Tsinghua UniversityCN1.2k papers
- Institute of High Energy PhysicsCN1k papers
- Centre National de la Recherche ScientifiqueFR1k papers
- Istituto Nazionale di Fisica Nucleare, Laboratori Nazionali di FrascatiIT987 papers
- University of Science and Technology of ChinaCN911 papers
- Istituto Nazionale di Fisica Nucleare, Sezione di PerugiaIT814 papers
- Beihang UniversityCN801 papers
- University of Maryland, College ParkUS776 papers
- National Centre for Nuclear ResearchPL769 papers
21 papers · 2 filters
UTDesign: A Unified Framework for Stylized Text Editing and Generation in Graphic Design Images
Yiming Zhao, Yuanpeng Gao, Yuxuan Luo +4
AI-assisted graphic design has emerged as a powerful tool for automating the creation and editing of design elements such as posters, banners, and advertisements. While diffusion-b…
High Dimensional Data Decomposition for Anomaly Detection of Textured Images
Ji Song, Xing Wang, Jianguo Wu +1
In the realm of diverse high-dimensional data, images play a significant role across various processes of manufacturing systems where efficient image anomaly detection has emerged…
Towards Deeper Emotional Reflection: Crafting Affective Image Filters with Generative Priors
Peixuan Zhang, Shuchen Weng, Jiajun Tang +2
Social media platforms enable users to express emotions by posting text with accompanying images. In this paper, we propose the Affective Image Filter (AIF) task, which aims to ref…
Stitch and Tell: A Structured Multimodal Data Augmentation Method for Spatial Understanding
Hang Yin, Xiaomin He, PeiWen Yuan +5
Existing vision-language models often suffer from spatial hallucinations, i.e., generating incorrect descriptions about the relative positions of objects in an image. We argue that…
Towards Reliable Human Evaluations in Gesture Generation: Insights from a Community-Driven State-of-the-Art Benchmark
Rajmund Nagy, Hendric Voss, Thanh Hoang-Minh +18
We review human evaluation practices in automatic, speech-driven 3D gesture generation and find a lack of standardisation and frequent use of flawed experimental setups. This leads…
RT-DETRv4: Painlessly Furthering Real-Time Object Detection with Vision Foundation Models
Zijun Liao, Yian Zhao, Xin Shan +5
Real-time object detection has achieved substantial progress through meticulously designed architectures and optimization strategies. However, the pursuit of high-speed inference v…