Walker: Self-supervised Multiple Object Tracking by Walking on Temporal Appearance Graphs

Segu, Mattia; Piccinelli, Luigi; Li, Siyuan; Van Gool, Luc; Yu, Fisher; Schiele, Bernt

Computer Science > Computer Vision and Pattern Recognition

arXiv:2409.17221 (cs)

[Submitted on 25 Sep 2024]

Title:Walker: Self-supervised Multiple Object Tracking by Walking on Temporal Appearance Graphs

Authors:Mattia Segu, Luigi Piccinelli, Siyuan Li, Luc Van Gool, Fisher Yu, Bernt Schiele

View PDF

Abstract:The supervision of state-of-the-art multiple object tracking (MOT) methods requires enormous annotation efforts to provide bounding boxes for all frames of all videos, and instance IDs to associate them through time. To this end, we introduce Walker, the first self-supervised tracker that learns from videos with sparse bounding box annotations, and no tracking labels. First, we design a quasi-dense temporal object appearance graph, and propose a novel multi-positive contrastive objective to optimize random walks on the graph and learn instance similarities. Then, we introduce an algorithm to enforce mutually-exclusive connective properties across instances in the graph, optimizing the learned topology for MOT. At inference time, we propose to associate detected instances to tracklets based on the max-likelihood transition state under motion-constrained bi-directional walks. Walker is the first self-supervised tracker to achieve competitive performance on MOT17, DanceTrack, and BDD100K. Remarkably, our proposal outperforms the previous self-supervised trackers even when drastically reducing the annotation requirements by up to 400x.

Comments:	ECCV 2024
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2409.17221 [cs.CV]
	(or arXiv:2409.17221v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2409.17221

Submission history

From: Mattia Segu [view email]
[v1] Wed, 25 Sep 2024 18:00:00 UTC (19,626 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Walker: Self-supervised Multiple Object Tracking by Walking on Temporal Appearance Graphs

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Walker: Self-supervised Multiple Object Tracking by Walking on Temporal Appearance Graphs

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators