Detect to track and track to detect

Feichtenhofer, C; Pinz, A; Zisserman, A

AI Collection

Conference item

Detect to track and track to detect

Abstract:: Recent approaches for high accuracy detection and tracking of object categories in video consist of complex multistage solutions that become more cumbersome each year. In this paper we propose a ConvNet architecture that jointly performs detection and tracking, solving the task in a simple and effective way. Our contributions are threefold: (i) we set up a ConvNet architecture for simultaneous detection and tracking, using a multi-task objective for frame-based object detection and across-frame track regression; (ii) we introduce correlation features that represent object co-occurrences across time to aid the ConvNet during tracking; and (iii) we link the frame level detections based on our across-frame tracklets to produce high accuracy detections at the video level. Our ConvNet architecture for spatiotemporal object detection is evaluated on the large-scale ImageNet VID dataset where it achieves state-of-the-art results. Our approach provides better single model performance than the winning method of the last ImageNet challenge while being conceptually much simpler. Finally, we show that by increasing the temporal stride we can dramatically increase the tracker speed.

Publication status:: Published

Peer review status:: Peer reviewed

Actions

Email

Email this record

Send the bibliographic details of this record to your email address.

Your Email
Please enter the email address that the record information will be sent to.

-
Your message (optional)
Please add any additional information to be included within the email.
Share
Cite

Cite this record

APA Style

Feichtenhofer, C., Pinz, A., & Zisserman, A. (2017). Detect to track and track to detect. 3057–3065.

MLA Style

Feichtenhofer, C, et al. “Detect to Track and Track to Detect.” 2017, pp. 3057–65.

Chicago Style

Feichtenhofer, C, A Pinz, and A Zisserman. 2017. “Detect to Track and Track to Detect.” 3057–65.
Print

Access Document

Files:: feichtenhofer17.pdf

(Preview, Accepted manuscript, pdf, 5.9MB, Terms of use)

Publisher copy:: 10.1109/ICCV.2017.330

Authors

+ Feichtenhofer, C More by this author

Role:: Author

+ Pinz, A More by this author

Role:: Author

+ Zisserman, A More by this author

Institution:: University of Oxford
Division:: MPLS
Department:: Engineering Science
Oxford college:: Brasenose College
Role:: Author
ORCID:: 0000-0002-8945-8573

+ Engineering and Physical Sciences Research Council More from this funder

Grant:: EP/M013774/1

+ Austrian Academy of Sciences More from this funder

Funding agency for:: Feichtenhofer, C

Publisher:: Institute of Electrical and Electronics Engineers
Host title:: IEEE International Conference on Computer Vision 2017
Journal:: IEEE International Conference on Computer Vision 2017 More from this journal
Pages:: 3057-3065
Publication date:: 2017-12-25
Acceptance date:: 2017-11-08
DOI:: 10.1109/ICCV.2017.330
ISSN:: 978-1-5386-1033-6, 1550-5499
ISBN:: 9781538610329

Pubs id:: pubs:821116
UUID:: uuid:638773c5-b132-4d89-ad99-1ac49a3c878c
Local pid:: pubs:821116
Source identifiers:: 821116
Deposit date:: 2018-08-17
ARK identifier:: ark:/29072/ora_638773c5b1324d89ad991ac49a3c878c

Terms of use

Copyright holder:: IEEE
Notes:: © 2017 IEEE. This is the author accepted manuscript following peer review version of the article. The final version is available online from IEEE at: 10.1109/ICCV.2017.330

Licence:: Terms and Conditions of Use for Oxford University Research Archive

Views and Downloads

About views and downloads

If you are the owner of this record, you can report an update to it here: Report update to this record

Conference item

Detect to track and track to detect

Actions

Access Document

Authors

Terms of use

Views and Downloads

Altmetrics

Dimensions

Conference item

Detect to track and track to detect

Actions

Access Document

Authors

Funding

Bibliographic Details

Item Description

Terms of use

Metrics

Views and Downloads

Altmetrics

Dimensions