Inferring Temporal Compositions of Actions Using Probabilistic Automata

Cruz, Rodrigo Santa; Cherian, Anoop; Fernando, Basura; Campbell, Dylan; Gould, Stephen

Computer Science > Computer Vision and Pattern Recognition

arXiv:2004.13217 (cs)

[Submitted on 28 Apr 2020]

Title:Inferring Temporal Compositions of Actions Using Probabilistic Automata

Authors:Rodrigo Santa Cruz, Anoop Cherian, Basura Fernando, Dylan Campbell, Stephen Gould

View PDF

Abstract:This paper presents a framework to recognize temporal compositions of atomic actions in videos. Specifically, we propose to express temporal compositions of actions as semantic regular expressions and derive an inference framework using probabilistic automata to recognize complex actions as satisfying these expressions on the input video features. Our approach is different from existing works that either predict long-range complex activities as unordered sets of atomic actions, or retrieve videos using natural language sentences. Instead, the proposed approach allows recognizing complex fine-grained activities using only pretrained action classifiers, without requiring any additional data, annotations or neural network training. To evaluate the potential of our approach, we provide experiments on synthetic datasets and challenging real action recognition datasets, such as MultiTHUMOS and Charades. We conclude that the proposed approach can extend state-of-the-art primitive action classifiers to vastly more complex activities without large performance degradation.

Comments:	Accepted in Workshop on Compositionality in Computer Vision at CVPR, 2020
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2004.13217 [cs.CV]
	(or arXiv:2004.13217v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2004.13217

Submission history

From: Rodrigo Santa Cruz [view email]
[v1] Tue, 28 Apr 2020 00:15:26 UTC (1,034 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Inferring Temporal Compositions of Actions Using Probabilistic Automata

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Inferring Temporal Compositions of Actions Using Probabilistic Automata

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators