3 papers
cs.CV2026
X-MULTI: VLM-based Imaging Factor Disentanglement for Factor-Aware Image Synthesis
Sonali Godavarthy, Matthias Neuwirth-Trapp, Tim-Felix Faasch +4
Imaging factor disentanglement in text-to-image generation aims to independently control image acquisition properties such as types of camera lenses, sensor types, viewpoints, and…
cs.CV2026
TASE: Truncation-Aware Semantic Embeddings for 3D Scene Understanding and Editing
Tim-Felix Faasch, Jochen Kall, Lucas Nunes +2
High-fidelity semantic 3D scene representations are crucial for numerous applications, including robotics, autonomous driving, and simulation. Beyond this, the ability to edit such…
cs.CV2026
MULTI: Disentangling Camera Lens, Sensor, View, and Domain for Novel Image Generation
Sonali Godavarthy, Matthias Neuwirth-Trapp, Tim-Felix Faasch +3
Recent text-to-image models produce high-quality images, yet text ambiguity hinders precise control when specific styles or objects are required. There have been a number of recent…