TRCKFRMER_PR: "TrackFormer: Multi-Object Tracking with Transformers" with private detections

MOT20-04


Benchmark:

MOT17 | MOT20 |

Short name:

TRCKFRMER_PR

Detector:

Private

Description:

The challenging task of multi-object tracking (MOT) requires simultaneous reasoning about track initialization, identity, and spatiotemporal trajectories. We formulate this task as a frame-to-frame set prediction problem and introduce TrackFormer, an end-to-end MOT approach based on an encoder-decoder Transformer architecture. Our model achieves data association between frames via attention by evolving a set of track predictions through a video sequence. The Transformer decoder initializes new tracks from static object queries and autoregressively follows existing tracks in space and time with the new concept of identity preserving track queries. Both decoder query types benefit from self- and encoder-decoder attention on global frame-level features, thereby omitting any additional graph optimization and matching or modeling of motion and appearance. TrackFormer represents a new tracking-by-attention paradigm and yields state-of-the-art performance on the task of multi-object tracking (MOT17) and segmentation (MOTS20).

Reference:

T. Meinhardt, A. Kirillov, L. Leal-Taixe, C. Feichtenhofer. TrackFormer: Multi-Object Tracking with Transformers. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2022.

Last submitted:

April 29, 2022 (1 year ago)

Published:

April 29, 2022 at 11:04:18 CET

Submissions:

1

Open source:

Yes

Hardware:

7 x 32 GB GPUs

Runtime:

5.7 Hz

Benchmark performance:

Sequence MOTA IDF1 HOTA MT ML FP FN Rcll Prcn AssA DetA AssRe AssPr DetRe DetPr LocA FAF ID Sw. Frag
MOT2068.665.754.7666 (53.6)181 (14.6)20,348140,37372.994.953.056.757.478.560.879.283.74.51,532 (0.0)2,474 (0.0)

Detailed performance:

Sequence MOTA IDF1 HOTA MT ML FP FN Rcll Prcn AssA DetA AssRe AssPr DetRe DetPr LocA FAF ID Sw. Frag
MOT20-0482.775.663.2490289,63937,16586.496.158.668.363.480.273.181.384.54.6566941
MOT20-0655.953.543.996725,58252,44060.593.542.245.846.074.149.175.981.75.5545842
MOT20-0756.259.049.2412054713,85658.197.250.248.453.383.550.684.786.90.992144
MOT20-0846.048.338.939614,58036,91252.489.939.738.543.471.841.971.981.15.7329547

Raw data:


TRCKFRMER_PR