Back to Search Start Over

TinyHD: Efficient video saliency prediction with heterogeneous decoders using hierarchical maps distillation

Authors :
Feiyan Hu
Simone Palazzo
Federica Proietto Salanitri
Giovanni Bellitto
Morteza Moradi
Concetto Spampinato
Kevin McGuinness
Source :
Hu, Feiyan ORCID: 0000-0001-7451-6438 , Palazzo, Simone ORCID: 0000-0002-2441-0982 , Proietto Salanitri, Federica ORCID: 0000-0002-6122-4249 , Bellitto, Giovanni, Moradi, Morteza, Spampinato, Concetto ORCID: 0000-0001-6653-2577 and McGuinness, Kevin ORCID: 0000-0003-1336-6477 (2022) TinyHD: Efficient video saliency prediction with heterogeneous decoders using hierarchical maps distillation. In: IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2023, 3-7 Jan 2023, Waikoloa, Hawaii. (In Press)
Publication Year :
2022
Publisher :
IEEE, 2022.

Abstract

Video saliency prediction has recently attracted attention of the research community, as it is an upstream task for several practical applications. However, current solutions are particularly computationally demanding, especially due to the wide usage of spatio-temporal 3D convolutions. We observe that, while different model architectures achieve similar performance on benchmarks, visual variations between predicted saliency maps are still significant. Inspired by this intuition, we propose a lightweight model that employs multiple simple heterogeneous decoders and adopts several practical approaches to improve accuracy while keeping computational costs low, such as hierarchical multi-map knowledge distillation, multi-output saliency prediction, unlabeled auxiliary datasets and channel reduction with teacher assistant supervision. Our approach achieves saliency prediction accuracy on par or better than state-of-the-art methods on DFH1K, UCF-Sports and Hollywood2 benchmarks, while enhancing significantly the efficiency of the model. Code is on https://github.com/feiyanhu/tinyHD<br />Comment: WACV2023

Details

Language :
English
Database :
OpenAIRE
Journal :
Hu, Feiyan ORCID: 0000-0001-7451-6438 <https://orcid.org/0000-0001-7451-6438>, Palazzo, Simone ORCID: 0000-0002-2441-0982 <https://orcid.org/0000-0002-2441-0982>, Proietto Salanitri, Federica ORCID: 0000-0002-6122-4249 <https://orcid.org/0000-0002-6122-4249>, Bellitto, Giovanni, Moradi, Morteza, Spampinato, Concetto ORCID: 0000-0001-6653-2577 <https://orcid.org/0000-0001-6653-2577> and McGuinness, Kevin ORCID: 0000-0003-1336-6477 <https://orcid.org/0000-0003-1336-6477> (2022) TinyHD: Efficient video saliency prediction with heterogeneous decoders using hierarchical maps distillation. In: IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2023, 3-7 Jan 2023, Waikoloa, Hawaii. (In Press)
Accession number :
edsair.doi.dedup.....09fd8ffb16cafef6baa263aa0fc4265e