End-to-End Representation Learning for Correlation Filter Based Tracking
Abstract
The Correlation Filter is an algorithm that trains a linear template to discriminate between images and their translations. It is well suited to object tracking because its formulation in the Fourier domain provides a fast solution, enabling the detector to be re-trained once per frame. Previous works that use the Correlation Filter, however, have adopted features that were either manually designed or trained for a different task. This work is the first to overcome this limitation by interpreting the Correlation Filter learner, which has a closed-form solution, as a differentiable layer in a deep neural network. This enables learning deep features that are tightly coupled to the Correlation Filter. Experiments illustrate that our method has the important practical benefit of allowing lightweight architectures to achieve state-of-the-art performance at high framerates.
Cite
Text
Valmadre et al. "End-to-End Representation Learning for Correlation Filter Based Tracking." Conference on Computer Vision and Pattern Recognition, 2017. doi:10.1109/CVPR.2017.531Markdown
[Valmadre et al. "End-to-End Representation Learning for Correlation Filter Based Tracking." Conference on Computer Vision and Pattern Recognition, 2017.](https://mlanthology.org/cvpr/2017/valmadre2017cvpr-endtoend/) doi:10.1109/CVPR.2017.531BibTeX
@inproceedings{valmadre2017cvpr-endtoend,
title = {{End-to-End Representation Learning for Correlation Filter Based Tracking}},
author = {Valmadre, Jack and Bertinetto, Luca and Henriques, Joao and Vedaldi, Andrea and Torr, Philip H. S.},
booktitle = {Conference on Computer Vision and Pattern Recognition},
year = {2017},
doi = {10.1109/CVPR.2017.531},
url = {https://mlanthology.org/cvpr/2017/valmadre2017cvpr-endtoend/}
}