OpenPifPaf: Composite Fields for Semantic Keypoint Detection and Spatio-Temporal Association

Kreiss, Sven; Bertoni, Lorenzo; Alahi, Alexandre

Computer Science > Computer Vision and Pattern Recognition

arXiv:2103.02440 (cs)

[Submitted on 3 Mar 2021 (v1), last revised 21 Sep 2021 (this version, v2)]

Title:OpenPifPaf: Composite Fields for Semantic Keypoint Detection and Spatio-Temporal Association

Authors:Sven Kreiss, Lorenzo Bertoni, Alexandre Alahi

View PDF

Abstract:Many image-based perception tasks can be formulated as detecting, associating and tracking semantic keypoints, e.g., human body pose estimation and tracking. In this work, we present a general framework that jointly detects and forms spatio-temporal keypoint associations in a single stage, making this the first real-time pose detection and tracking algorithm. We present a generic neural network architecture that uses Composite Fields to detect and construct a spatio-temporal pose which is a single, connected graph whose nodes are the semantic keypoints (e.g., a person's body joints) in multiple frames. For the temporal associations, we introduce the Temporal Composite Association Field (TCAF) which requires an extended network architecture and training method beyond previous Composite Fields. Our experiments show competitive accuracy while being an order of magnitude faster on multiple publicly available datasets such as COCO, CrowdPose and the PoseTrack 2017 and 2018 datasets. We also show that our method generalizes to any class of semantic keypoints such as car and animal parts to provide a holistic perception framework that is well suited for urban mobility such as self-driving cars and delivery robots.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2103.02440 [cs.CV]
	(or arXiv:2103.02440v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2103.02440

Submission history

From: Sven Kreiss [view email]
[v1] Wed, 3 Mar 2021 14:44:14 UTC (14,908 KB)
[v2] Tue, 21 Sep 2021 09:35:40 UTC (15,103 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:OpenPifPaf: Composite Fields for Semantic Keypoint Detection and Spatio-Temporal Association

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:OpenPifPaf: Composite Fields for Semantic Keypoint Detection and Spatio-Temporal Association

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators