2 papers
cs.CV2026
Image-Space Rule Discovery
Misora Sugiyama, Toya Oyama, Hirokatsu Kataoka
Can image-editing models discover visual rules in image space and complete problem-solving end-to-end? We tackle this question in the spirit of a human worksheet test (e.g., an IQ…
cs.CV2025
Simple Visual Artifact Detection in Sora-Generated Videos
Misora Sugiyama, Hirokatsu Kataoka
The December 2024 release of OpenAI's Sora, a powerful video generation model driven by natural language prompts, highlights a growing convergence between large language models (LL…