Selective Feature Compression for Efficient Activity Recognition Inference

Liu, Chunhui; Li, Xinyu; Chen, Hao; Modolo, Davide; Tighe, Joseph

Full-text links:

Download:

Current browse context:

cs.CV

< prev | next >

new | recent | 2104

Change to browse by:

Computer Science > Computer Vision and Pattern Recognition

Title: Selective Feature Compression for Efficient Activity Recognition Inference

Authors: Chunhui Liu, Xinyu Li, Hao Chen, Davide Modolo, Joseph Tighe

(Submitted on 1 Apr 2021 (v1), last revised 29 Jul 2021 (this version, v2))

Abstract: Most action recognition solutions rely on dense sampling to precisely cover the informative temporal clip. Extensively searching temporal region is expensive for a real-world application. In this work, we focus on improving the inference efficiency of current action recognition backbones on trimmed videos, and illustrate that one action model can also cover then informative region by dropping non-informative features. We present Selective Feature Compression (SFC), an action recognition inference strategy that greatly increase model inference efficiency without any accuracy compromise. Differently from previous works that compress kernel sizes and decrease the channel dimension, we propose to compress feature flow at spatio-temporal dimension without changing any backbone parameters. Our experiments on Kinetics-400, UCF101 and ActivityNet show that SFC is able to reduce inference speed by 6-7x and memory usage by 5-6x compared with the commonly used 30 crops dense sampling procedure, while also slightly improving Top1 Accuracy. We thoroughly quantitatively and qualitatively evaluate SFC and all its components and show how does SFC learn to attend to important video regions and to drop temporal features that are uninformative for the task of action recognition.

Comments:	Accepted by ICCV 2021
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2104.00179 [cs.CV]
	(or arXiv:2104.00179v2 [cs.CV] for this version)

Submission history

From: Chunhui Liu [view email]
[v1] Thu, 1 Apr 2021 00:54:51 GMT (28155kb,D)
[v2] Thu, 29 Jul 2021 10:59:15 GMT (28155kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2104.00179

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computer Vision and Pattern Recognition

Title: Selective Feature Compression for Efficient Activity Recognition Inference

Submission history