Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
SATURN: Symbolic Spatial Reasoning for Multi-Perspective Grounding
Danial Kamali, Tanawan Premsri, Shreya Rajpal +3
Vision-Language Models (VLMs) remain unreliable when spatial reasoning requires composing relations whose meanings depend on frames of reference. Existing neuro-symbolic methods ma…
cs.CV2025
FoR-SALE: Frame of Reference-guided Spatial Adjustment in LLM-based Diffusion Editing
Tanawan Premsri, Parisa Kordjamshidi
Current text-to-image generation models, even state-of-the-art models, exhibit a significant performance gap when spatial expressions are described from non-camera perspectives. To…