3 papers
cs.CV2026
Listener-Rewarded Thinking in VLMs for Image Preferences
Alexander Gambashidze, Li Pengyi, Matvey Skripkin +5
Training robust and generalizable reward models for human visual preferences is essential for aligning text-to-image and text-to-video generative models with human intent. However,…
cs.CV2026
CoMa: Contextual Massing Generation with Vision-Language Models
Evgenii Maslov, Valentin Khrulkov, Anastasia Volkova +3
The conceptual design phase in architecture and urban planning, particularly building massing, is complex and heavily reliant on designer intuition and manual effort. To address th…
cs.AI2025
Multi-Agent GraphRAG: A Text-to-Cypher Framework for Labeled Property Graphs
Anton Gusarov, Anastasia Volkova, Valentin Khrulkov +3
While Retrieval-Augmented Generation (RAG) methods commonly draw information from unstructured documents, the emerging paradigm of GraphRAG aims to leverage structured data such as…