works on

From the 1 of 9 linked papers with an AI index.

collaborators
Showing cs.CVShow all

8 papers · 1 filter

cs.CV2026

DVPSFormer: Efficient Online Depth-aware Video Panoptic Segmentation for Autonomous Driving

Yung-Hsu Yang, Luigi Piccinelli, Siyuan Li +8

DVPSFormer is an online architecture that jointly estimates metric depth, semantic segmentation, and instance trajectories for autonomous driving by using explicit scene discretiza…

cs.CV2024

SLAck: Semantic, Location, and Appearance Aware Open-Vocabulary Tracking

Siyuan Li, Lei Ke, Yung-Hsu Yang +4

Open-vocabulary Multiple Object Tracking (MOT) aims to generalize trackers to novel categories not in the training set. Currently, the best-performing methods are mainly based on p…

cs.CV2024

PIV3CAMS: a multi-camera dataset for multiple computer vision problems and its application to novel view-point synthesis

Sohyeong Kim, Martin Danelljan, Radu Timofte +2

The modern approaches for computer vision tasks significantly rely on machine learning, which requires a large number of quality images. While there is a plethora of image datasets…

cs.CV2024

Gaussian Grouping: Segment and Edit Anything in 3D Scenes

Mingqiao Ye, Martin Danelljan, Fisher Yu +1

The recent Gaussian Splatting achieves high-quality and real-time novel-view synthesis of the 3D scenes. However, it is solely concentrated on the appearance and geometry modeling,…

cs.CV2024

Matching Anything by Segmenting Anything

Siyuan Li, Lei Ke, Martin Danelljan +4

The robust association of the same objects across video frames in complex scenes is crucial for many applications, especially Multiple Object Tracking (MOT). Current methods predom…

cs.CV2024

Analyzing Local Representations of Self-supervised Vision Transformers

Ani Vanyan, Alvard Barseghyan, Hakob Tamazyan +3

In this paper, we present a comparative analysis of various self-supervised Vision Transformers (ViTs), focusing on their local representative power. Inspired by large language mod…