Advances in Artificial Intelligence: A Review for the Creative Industries
arXiv:2501.02725 · doi:10.1007/s10462-026-11494-w
Abstract
Artificial intelligence (AI) has undergone transformative advances since 2022, particularly through generative AI, large language models (LLMs), and diffusion models, fundamentally reshaping the creative industries. However, existing reviews have not comprehensively addressed these recent breakthroughs and their integrated impact across the creative production pipeline. This paper addresses this gap by providing a systematic review of AI technologies that have emerged or matured since our 2022 review, examining their applications across content creation, information analysis, post-production enhancement, compression, and quality assessment. We document how transformers, LLMs, diffusion models, and implicit neural representations have established new capabilities in text-to-image/video generation, real-time 3D reconstruction, and unified multi-task frameworks-shifting AI from support tool to core creative technology. Beyond technological advances, we analyze the trend toward unified AI frameworks that integrate multiple creative tasks, replacing task-specific solutions. We critically examine the evolving role of human-AI collaboration, where human oversight remains essential for creative direction and mitigating AI hallucinations. Finally, we identify emerging challenges including copyright concerns, bias mitigation, computational demands, and the need for robust regulatory frameworks. This review provides researchers and practitioners with a comprehensive understanding of current AI capabilities, limitations, and future trajectories in creative applications.
This is an updated review of our previous paper (see https://doi.org/10.1007/s10462-021-10039-7), and has been accepted by Artificial Intelligence Review journal
References in corpus (38)
- Instant Neural Graphics Primitives with a Multiresolution Hash Encoding
- Attention Mechanisms in Computer Vision: A Survey
- Vision Transformers for Single Image Dehazing
- A Survey on Large Language Model (LLM) Security and Privacy: The Good, the Bad, and the Ugly
- Artificial Intelligence in the Creative Industries: A Review
- CrossFuse: A Novel Cross Attention Mechanism based Infrared and Visible Image Fusion Approach
- Uncertainty-Aware Blind Image Quality Assessment in the Laboratory and Wild
- Large-Scale Study of Perceptual Video Quality
- Image Quality Assessment using Contrastive Learning
- Fast Dynamic Radiance Fields with Time-Aware Neural Voxels
- YouTube UGC Dataset for Video Compression Research
- SUNet: Swin Transformer UNet for Image Denoising
- DeepQTMT: A Deep Learning Approach for Fast QTMT-based CU Partition of Intra-mode VVC
- 3D Gaussian as a New Era: A Survey
- Video Transformers: A Survey
- Transformers and Large Language Models for Efficient Intrusion Detection Systems: A Comprehensive Survey
- A Design Space for Intelligent and Interactive Writing Assistants
- Diffusion Models, Image Super-Resolution And Everything: A Survey
- Diffusion Model-Based Image Editing: A Survey
- Rethinking Search: Making Domain Experts out of Dilettantes
- Sasha: Creative Goal-Oriented Reasoning in Smart Homes with Large Language Models
- Transformers in Single Object Tracking: An Experimental Survey
- CLE Diffusion: Controllable Light Enhancement Diffusion Model
- Subjective and Objective Quality Assessment of High Frame Rate Videos
- Transformer-based Image and Video Inpainting: Current Challenges and Future Directions
- Reviewing Intelligent Cinematography: AI research for camera-based video production
- RankDVQA: Deep VQA based on Ranking-inspired Hybrid Training
- UW-GS: Distractor-Aware 3D Gaussian Splatting for Enhanced Underwater Scene Reconstruction
- Cool-chic video: Learned video coding with 800 parameters
- BVI-AOM: A New Training Dataset for Deep Video Compression Optimization
- Immersive Video Compression using Implicit Neural Representations
- BVI-Artefact: An Artefact Detection Benchmark Dataset for Streamed Videos
- Content-Based Search for Deep Generative Models
- Accelerating Learnt Video Codecs with Gradient Decay and Layer-wise Distillation
- CleAR: Robust Context-Guided Generative Lighting Estimation for Mobile Augmented Reality
- RMT-BVQA: Recurrent Memory Transformer-based Blind Video Quality Assessment for Enhanced Video Content
- NTIRE 2025 Challenge on Image Super-Resolution (x4): Methods and Results
- DaBiT: Depth and Blur informed Transformer for Video Focal Deblurring