Context-aware multimodal AI navigates hidden pathways in five centuries of art evolution

Citations

WEB OF SCIENCE

0
Citations

SCOPUS

0

초록

The rise of multimodal generative AI transforms the intersection of technology and art, offering richer insights into large-scale artworks. While significant research has focused on their creative potential, their ability to represent artworks in latent spaces remains underexamined. We use generative AI, specifically Stable Diffusion, to analyze 500 y of Western paintings by extracting two types of latent information with the model: formal aspects (e.g., colors) and contextual aspects (e.g., subjects). Our findings reveal that contextual information exhibits stronger vector alignment and orientation with conventional artistic periods, styles, and individual artists than formal elements. Also, we show how artistic expression aligns with historical shifts using contextual keywords extracted from paintings. Our generative experiment, infusing prospective contexts into historical artworks, validates this vector alignment and orientation by synthesizing artworks consistent with the stylistic patterns of target periods. This study demonstrates how multimodal AI expands traditional formal analysis by integrating temporal, cultural, and historical contexts to quantify the latent structure of cultural knowledge.

키워드

art historycontextual informationmultimodal AIcultural evolutioncomputational humanitiesHISTORY
제목
Context-aware multimodal AI navigates hidden pathways in five centuries of art evolution
저자
Kim, JinLee, ByunghweeYou, TaekhoYun, Jinhyuk
DOI
10.1073/pnas.2517969123
발행일
2026-07
유형
Article
저널명
Proceedings of the National Academy of Sciences of the United States of America
123
30