For a semiotic AI: Bridging computer vision and visual semiotics for computational observation of large scale facial image archives
arXiv:2407.03268 · doi:10.1016/j.cviu.2024.104187
Abstract
Social networks are creating a digital world in which the cognitive, emotional, and pragmatic value of the imagery of human faces and bodies is arguably changing. However, researchers in the digital humanities are often ill-equipped to study these phenomena at scale. This work presents FRESCO (Face Representation in E-Societies through Computational Observation), a framework designed to explore the socio-cultural implications of images on social media platforms at scale. FRESCO deconstructs images into numerical and categorical variables using state-of-the-art computer vision techniques, aligning with the principles of visual semiotics. The framework analyzes images across three levels: the plastic level, encompassing fundamental visual features like lines and colors; the figurative level, representing specific entities or concepts; and the enunciation level, which focuses particularly on constructing the point of view of the spectator and observer. These levels are analyzed to discern deeper narrative layers within the imagery. Experimental validation confirms the reliability and utility of FRESCO, and we assess its consistency and precision across two public datasets. Subsequently, we introduce the FRESCO score, a metric derived from the framework's output that serves as a reliable measure of similarity in image content.
Accepted at CVIU journal 2024
References in corpus (9)
- A Comprehensive Survey on Pretrained Foundation Models: A History from BERT to ChatGPT
- Exploring the Limits of Out-of-Distribution Detection
- What your Facebook Profile Picture Reveals about your Personality
- Prismer: A Vision-Language Model with Multi-Task Experts
- From colouring-in to pointillism: revisiting semantic segmentation supervision
- Latent Diffusion Models for Attribute-Preserving Image Anonymization
- Automatic Image Content Extraction: Operationalizing Machine Learning in Humanistic Photographic Studies of Large Visual Archives
- Seeing the Intangible: Survey of Image Classification into High-Level and Abstract Categories
- 3DGazeNet: Generalizing Gaze Estimation with Weak-Supervision from Synthetic Views