Temporal Extension of Scale Pyramid and Spatial Pyramid Matching for Action Recognition

Lan, Zhenzhong; Li, Xuanchong; Hauptmann, Alexandar G.

Computer Science > Computer Vision and Pattern Recognition

arXiv:1408.7071 (cs)

[Submitted on 29 Aug 2014]

Title:Temporal Extension of Scale Pyramid and Spatial Pyramid Matching for Action Recognition

Authors:Zhenzhong Lan, Xuanchong Li, Alexandar G. Hauptmann

View PDF

Abstract:Historically, researchers in the field have spent a great deal of effort to create image representations that have scale invariance and retain spatial location information. This paper proposes to encode equivalent temporal characteristics in video representations for action recognition. To achieve temporal scale invariance, we develop a method called temporal scale pyramid (TSP). To encode temporal information, we present and compare two methods called temporal extension descriptor (TED) and temporal division pyramid (TDP) . Our purpose is to suggest solutions for matching complex actions that have large variation in velocity and appearance, which is missing from most current action representations. The experimental results on four benchmark datasets, UCF50, HMDB51, Hollywood2 and Olympic Sports, support our approach and significantly outperform state-of-the-art methods. Most noticeably, we achieve 65.0% mean accuracy and 68.2% mean average precision on the challenging HMDB51 and Hollywood2 datasets which constitutes an absolute improvement over the state-of-the-art by 7.8% and 3.9%, respectively.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1408.7071 [cs.CV]
	(or arXiv:1408.7071v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1408.7071

Submission history

From: Zhenzhong Lan [view email]
[v1] Fri, 29 Aug 2014 17:05:29 UTC (7,299 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CV

< prev | next >

new | recent | 2014-08

Change to browse by:

References & Citations

DBLP - CS Bibliography

listing | bibtex

Zhen-Zhong Lan
Xuanchong Li
Alexander G. Hauptmann

export BibTeX citation

Computer Science > Computer Vision and Pattern Recognition

Title:Temporal Extension of Scale Pyramid and Spatial Pyramid Matching for Action Recognition

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Temporal Extension of Scale Pyramid and Spatial Pyramid Matching for Action Recognition

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators