Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
World in a Frame: Understanding Culture Mixing as a New Challenge for Vision-Language Models
Eunsu Kim, Junyeong Park, Na Min An +9
In a globalized world, cultural elements from diverse origins frequently appear together within a single visual scene. We refer to these as culture mixing scenarios, yet how Large…
cs.CV2024
Semi-Truths: A Large-Scale Dataset of AI-Augmented Images for Evaluating Robustness of AI-Generated Image detectors
Anisha Pal, Julia Kruk, Mansi Phute +4
Text-to-image diffusion models have impactful applications in art, design, and entertainment, yet these technologies also pose significant risks by enabling the creation and dissem…
cs.CV2019
Integrating Text and Image: Determining Multimodal Document Intent in Instagram Posts
Julia Kruk, Jonah Lubin, Karan Sikka +3
Computing author intent from multimodal data like Instagram posts requires modeling a complex relationship between text and image. For example, a caption might evoke an ironic cont…