3 citations · 4 across the 9 of their papers we have counts for
19 papers
Multimodal Floorplan Encoding: Learning Dense Modality-Invariant Representations
Xavier Anadón, Rémi Pautrat, Rui Wang
Floorplans arise in many forms, from vector CAD drawings to raster renderings and sensor-derived density maps. This heterogeneity makes it difficult to build learning systems that…
Pixel-wise Planarity for High-Precision Monocular Plane Segmentation
Ahmetcan Yavuz, Alpay Ozkan, Rémi Pautrat +2
Plane segmentation from a single RGB image remains challenging due to imprecise region grouping and geometrically inconsistent supervision, often leading to over-segmentation and f…
Stable and Scalable Bundle Adjustment of Holistic 3D Structures
Shaohui Liu, Rémi Pautrat, Daniel Barath +3
Bundle Adjustment (BA) is a cornerstone of 3D computer vision and has benefited from decades of advances in sparse optimization and numerical methods. It was originally developed f…
Unified and Efficient Point-Line Local Features
François Costa, Raphael Kreft, Eckhard Goedeke +6
Multi-view computer vision pipelines typically rely on accurate sparse keypoints and robust descriptors. While incorporating line features has shown clear benefits for matching and…
PolyLayout: Multi-room Manhattan Layout Estimation
Gustav Hanning, Shaohui Liu, Rémi Pautrat +3
Estimating room layouts from multi-view imagery is a core task for indoor scene understanding. Existing methods are typically limited either by poor generalization to new datasets…
CoPE-VideoLM: Leveraging Codec Primitives For Efficient Video Language Modeling
Sayan Deb Sarkar, Rémi Pautrat, Ondrej Miksik +4
Video Language Models (VideoLMs) enable AI systems to understand temporal dynamics in videos. To fit within the maximum context window constraint, current methods use keyframe samp…