Hallucinating Pose-Compatible Scenes

Brooks, Tim; Efros, Alexei A.

Computer Science > Computer Vision and Pattern Recognition

arXiv:2112.06909 (cs)

[Submitted on 13 Dec 2021 (v1), last revised 1 Oct 2022 (this version, v2)]

Title:Hallucinating Pose-Compatible Scenes

Authors:Tim Brooks, Alexei A. Efros

View PDF

Abstract:What does human pose tell us about a scene? We propose a task to answer this question: given human pose as input, hallucinate a compatible scene. Subtle cues captured by human pose -- action semantics, environment affordances, object interactions -- provide surprising insight into which scenes are compatible. We present a large-scale generative adversarial network for pose-conditioned scene generation. We significantly scale the size and complexity of training data, curating a massive meta-dataset containing over 19 million frames of humans in everyday environments. We double the capacity of our model with respect to StyleGAN2 to handle such complex data, and design a pose conditioning mechanism that drives our model to learn the nuanced relationship between pose and scene. We leverage our trained model for various applications: hallucinating pose-compatible scene(s) with or without humans, visualizing incompatible scenes and poses, placing a person from one generated image into another scene, and animating pose. Our model produces diverse samples and outperforms pose-conditioned StyleGAN2 and Pix2Pix/Pix2PixHD baselines in terms of accurate human placement (percent of correct keypoints) and quality (Frechet inception distance).

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2112.06909 [cs.CV]
	(or arXiv:2112.06909v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2112.06909

Submission history

From: Tim Brooks [view email]
[v1] Mon, 13 Dec 2021 18:59:26 UTC (33,171 KB)
[v2] Sat, 1 Oct 2022 00:31:56 UTC (8,964 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Hallucinating Pose-Compatible Scenes

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Hallucinating Pose-Compatible Scenes

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators