Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
SceneGraphVLM: Dynamic Scene Graph Generation from Video with Vision-Language Models
Vladislav Makarov, Mark Gizetdinov, Dmitry Yudin
Scene graph generation provides a compact structured representation for visual perception, but accurate and fast graph prediction from images and videos remains challenging. Recent…
cs.CV2025
RCDINO: Enhancing Radar-Camera 3D Object Detection with DINOv2 Semantic Features
Olga Matykina, Dmitry Yudin
Three-dimensional object detection is essential for autonomous driving and robotics, relying on effective fusion of multimodal data from cameras and radar. This work proposes RCDIN…
cs.CV2024
uSF: Learning Neural Semantic Field with Uncertainty
Vsevolod Skorokhodov, Darya Drozdova, Dmitry Yudin
Recently, there has been an increased interest in NeRF methods which reconstruct differentiable representation of three-dimensional scenes. One of the main limitations of such meth…